Llama 3.3 70B Instruct

Meta · meta-llama/llama-3.3-70b-instruct

← Back to leaderboard

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

Llama 3.3 70B Instruct ranks #104 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.21 blended per million tokens, it is cheaper than 80% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).

open weightstexttext->text

Context

Max context: 131072
Max output: 16384

Pricing

Input / 1M: 0.10
Output / 1M: 0.32
Blend / 1M: 0.21

Quality

Quality index:

Provider

Provider: Meta
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.21per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.21
10,000,000$2.10
100,000,000$21.00

Related models

Same lab, similar price, or similar quality — useful next comparisons.

Closest API price