Mistral Nemo
Mistral · mistralai/mistral-nemo
A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...
Mistral Nemo ranks #50 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.02 blended per million tokens, it is cheaper than 97% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 16384
Pricing
Output / 1M: 0.03
Blend / 1M: 0.02
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.02per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.02 |
| 10,000,000 | $0.24 |
| 100,000,000 | $2.45 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Mistral
- Mistral Small 3$0.07 / 1M
- Ministral 3 3B 2512$0.10 / 1M
- Mistral Small 3.2 24B$0.14 / 1M
- Ministral 3 8B 2512$0.15 / 1M
- Ministral 3 14B 2512$0.20 / 1M
Closest API price
- Ling 3.0 Flash$0.04 / 1M
- Llama 3 8B Lunaris$0.04 / 1M
- Ling 3.0 Flash Sante$0.00 / 1M
- Dots3-Note Preview$0.00 / 1M
- LFM2.5-2.6B$0.00 / 1M