Skip to content
Türkçe

DeepSeek · Models · Pricing and access

DeepSeek V4.1 Flash is out: 552 billion parameters, but only 8 billion active per token

Published: 4 sourcesTürkçe

It arrives with a new causal encoder-decoder architecture and is almost twice the size of its predecessor, yet the price fell. Peak-hour rates are $0.30 per million input tokens and $1.20 output. On the retired V4 Flash those were $0.44 and $1.32. The weights were released under an MIT license.

Sources

  1. DeepSeek, “DeepSeek-V4.1-Flash: Smarter, Faster, More Efficient”, launch announcement in API docs, (api-docs.deepseek.com)
  2. DeepSeek, “Models & Pricing”, API pricing page (current) (api-docs.deepseek.com)
  3. DeepSeek (Internet Archive kopyası), “Models & Pricing”, archived copy of the pricing page from September 9, 2026 (V4 Flash prices), (web.archive.org)
  4. Hugging Face (DeepSeek), “deepseek-ai/DeepSeek-V4.1-Flash”, model weights page, (huggingface.co)

About this story

This story was posted on Instagram by @jarrus.tech on Sept. 15, 2026.

Spotted an error in this story? [email protected] · Instagram

This story in Turkish: DeepSeek V4.1 Flash çıktı: 552 milyar parametre, ama token başına sadece 8 milyarı aktif

On the same topic