R1 Distill Llama 70B
DeepSeek · deepseek/deepseek-r1-distill-llama-70b
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...
R1 Distill Llama 70B ranks #183 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.80 blended per million tokens, it is cheaper than 57% of models with published API pricing. Its 8,192-token context window is 32.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 7372
Pricing
Output / 1M: 0.80
Blend / 1M: 0.80
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.80per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.80 |
| 10,000,000 | $8.00 |
| 100,000,000 | $80.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from DeepSeek
- DeepSeek V4 Pro 0423$1.09 / 1M · Q 88.2
- DeepSeek V4 Pro 0813$2.24 / 1M · Q 88.2
- DeepSeek V4 Flash 0423$0.12 / 1M · Q 88.0
- DeepSeek V4 Flash 0731$0.12 / 1M · Q 88.0
- DeepSeek V3.2$0.33 / 1M
Closest API price
- Qwen3.7 Plus$0.80 / 1M · Q 78.7
- Qwen3.5-35B-A3B$0.78 / 1M
- Inkling Small$0.82 / 1M
- Perceptron Mk1$0.82 / 1M
- Qwen2.5 Coder 32B Instruct$0.83 / 1M