DeepSeek V4 Flash 0423
DeepSeek · deepseek/deepseek-v4-flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
DeepSeek V4 Flash 0423 ranks #10 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.13 blended per million tokens, it is cheaper than 89% of models with published API pricing. Its 1,024,000-token context window is 3.9× the catalog median (262,144 tokens). A quality index of 86.3 places it at #9 of 41 models with an Arena Elo signal.
Context
Max output: 384000
Pricing
Output / 1M: 0.18
Blend / 1M: 0.13
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.13per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.13 |
| 10,000,000 | $1.33 |
| 100,000,000 | $13.29 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from DeepSeek
- DeepSeek V4 Pro 0423$1.43 / 1M · Q 86.5
- DeepSeek V4 Pro 0813$2.10 / 1M · Q 86.5
- DeepSeek V4 Flash 0731$0.12 / 1M · Q 86.3
- DeepSeek V3.2$0.33 / 1M
- DeepSeek V3.2 Exp$0.34 / 1M
Closest API price
- Laguna S 2.1$0.14 / 1M
- Mistral Small 3.2 24B$0.14 / 1M
- Nemotron 3.5 Lightning$0.14 / 1M
- Granite 4.2 8B$0.13 / 1M
- Qwen3.5-9B$0.13 / 1M
Closest quality index
- DeepSeek V4 Flash 0731$0.12 / 1M · Q 86.3
- DeepSeek V4 Pro 0423$1.43 / 1M · Q 86.5
- DeepSeek V4 Pro 0813$2.10 / 1M · Q 86.5
- Gemini 3.7 Flash$2.25 / 1M · Q 86.7
- Qwen3.8 27B$1.71 / 1M · Q 87.4