DeepSeek · Models · Pricing and access
DeepSeek V4.1 Flash is out: 552 billion parameters, but only 8 billion active per token
It arrives with a new causal encoder-decoder architecture and is almost twice the size of its predecessor, yet the price fell. Peak-hour rates are $0.30 per million input tokens and $1.20 output. On the retired V4 Flash those were $0.44 and $1.32. The weights were released under an MIT license.
Sources
- DeepSeek, “DeepSeek-V4.1-Flash: Smarter, Faster, More Efficient”, launch announcement in API docs, (api-docs.deepseek.com)
- DeepSeek, “Models & Pricing”, API pricing page (current) (api-docs.deepseek.com)
- DeepSeek (Internet Archive kopyası), “Models & Pricing”, archived copy of the pricing page from September 9, 2026 (V4 Flash prices), (web.archive.org)
- Hugging Face (DeepSeek), “deepseek-ai/DeepSeek-V4.1-Flash”, model weights page, (huggingface.co)
About this story
This story was posted on Instagram by @jarrus.tech on Sept. 15, 2026.
Spotted an error in this story? [email protected] · Instagram
Short link: thejarrus.com/en/h3-13
This story in Turkish: DeepSeek V4.1 Flash çıktı: 552 milyar parametre, ama token başına sadece 8 milyarı aktif
On the same topic
Anthropic releases Claude Haiku 5.5 at 90% lower prices
Claude Haiku 5.5, released Oct. 7, costs 90% less than Haiku 4.5 for prompts up to 100,000 tokens. It leads on benchmarks but uses more tokens per task.
OpenAI is rolling out GPT-6 to everyone in ChatGPT
OpenAI began rolling out GPT-6 to everyone in ChatGPT on October 7. Its new Intelligent UI builds answers from text, visuals and interactive elements.
The open model war between Europe, the US and China heats up
Mistral unveiled Large 4 on October 6. The same week, Ecosia dropped Mistral, Reflection AI launched Beam and DeepSeek was reported near a $12 billion round.