DeepSeek V4.1 Flash
DeepSeek · deepseek/deepseek-v4.1-flash
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
DeepSeek V4.1 Flash ranks #94 of 359 models in the default AI Benchmark Hub order (quality first, then cost). At $0.38 blended per million tokens, it is cheaper than 72% of models with published API pricing. Its 1,048,576-token context window is 4.0× the catalog median (262,144 tokens).
Context
Max output: 384000
Pricing
Output / 1M: 0.60
Blend / 1M: 0.38
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.38per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.38 |
| 10,000,000 | $3.75 |
| 100,000,000 | $37.50 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from DeepSeek
- DeepSeek V4 Flash 0423$0.06 / 1M
- DeepSeek V4 Flash 0731$0.06 / 1M
- DeepSeek V3.2$0.33 / 1M
- DeepSeek V3.2 Exp$0.34 / 1M
- DeepSeek V4 Flash Vision Exp$0.43 / 1M
Closest API price
- Mistral Small 4$0.38 / 1M
- Solar Pro 3$0.38 / 1M
- gpt-oss-120b$0.38 / 1M
- Command R (08-2024)$0.38 / 1M
- GPT-4o-mini$0.38 / 1M