Llama 3.2 1B Instruct
Meta · meta-llama/llama-3.2-1b-instruct
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...
Llama 3.2 1B Instruct ranks #73 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.11 blended per million tokens, it is cheaper than 91% of models with published API pricing. Its 60,000-token context window is 4.4× smaller than the catalog median (262,144 tokens).
Context
Max output: 54000
Pricing
Output / 1M: 0.20
Blend / 1M: 0.11
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.11per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.11 |
| 10,000,000 | $1.14 |
| 100,000,000 | $11.40 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Meta
- Llama 3.1 8B Instruct$0.07 / 1M
- Muse Spark 1.3 Contributor$0.15 / 1M
- Muse Spark 1.2 Contributor$0.15 / 1M
- Llama Guard 4 12B$0.18 / 1M
- Llama 3.2 3B Instruct$0.19 / 1M
Closest API price
- Hy-MT2-1.8B$0.11 / 1M
- Ling 3.0 Flash Fin$0.12 / 1M
- Qwen3 30B A3B Instruct 2507$0.12 / 1M
- DeepSeek V4 Flash 0731$0.12 / 1M · Q 86.3
- Phi 4$0.11 / 1M