Llama 3.2 3B Instruct is 9.0× cheaper per million tokens (blended). Qwen3.8 27B has the larger context window (1,000,000 vs 131,072 tokens, 7.6×).

Llama 3.2 3B Instruct vs Qwen3.8 27B

Live catalog fields. Best-in-row is highlighted.

Llama 3.2 3B Instructopen

Meta · meta-llama__llama-3.2-3b-instruct

Qwen3.8 27Bopen

Qwen · qwen__qwen3.8-27b

Identity

Fieldmeta-llama__llama-3.2-3b-instructqwen__qwen3.8-27b
ProviderMetaQwen
Slugmeta-llama__llama-3.2-3b-instructqwen__qwen3.8-27b
Statuslivelive
Open weightsYesYes
License
HuggingFacemeta-llama/Llama-3.2-3B-InstructQwen/Qwen3.8-27B

Quality

Fieldmeta-llama__llama-3.2-3b-instructqwen__qwen3.8-27b
Quality index (0–100)87.4

Cost

Fieldmeta-llama__llama-3.2-3b-instructqwen__qwen3.8-27b
Input $ / 1M tokens$0.05$0.42
Cached input $ / 1M$0.09
Output $ / 1M tokens$0.33$3.00

Context

Fieldmeta-llama__llama-3.2-3b-instructqwen__qwen3.8-27b
Max context tokens131,0721,000,000
Max output tokens117,964131,072

Modalities

Fieldmeta-llama__llama-3.2-3b-instructqwen__qwen3.8-27b
Inputtexttext, image, video
Outputtexttext