Llama 4 Maverick

Meta · meta-llama/llama-4-maverick

← Back to leaderboard

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

Llama 4 Maverick ranks #136 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.45 blended per million tokens, it is cheaper than 71% of models with published API pricing. Its 128,000-token context window is 2.0× smaller than the catalog median (262,144 tokens).

open weightsimagetexttext+image->text

Context

Max context: 128000
Max output: 115200

Pricing

Input / 1M: 0.20
Output / 1M: 0.70
Blend / 1M: 0.45

Quality

Quality index:

Provider

Provider: Meta
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.45per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.45
10,000,000$4.48
100,000,000$44.80

Related models

Same lab, similar price, or similar quality — useful next comparisons.

Closest API price