Llama 3.2 1B Instruct is 15.0× cheaper per million tokens (blended). Qwen3.8 27B has the larger context window (1,000,000 vs 60,000 tokens, 16.7×).

Llama 3.2 1B Instruct vs Qwen3.8 27B

Live catalog fields. Best-in-row is highlighted.

Llama 3.2 1B Instructopen

Meta · meta-llama__llama-3.2-1b-instruct

Qwen3.8 27Bopen

Qwen · qwen__qwen3.8-27b

Identity

Fieldmeta-llama__llama-3.2-1b-instructqwen__qwen3.8-27b
ProviderMetaQwen
Slugmeta-llama__llama-3.2-1b-instructqwen__qwen3.8-27b
Statuslivelive
Open weightsYesYes
License
HuggingFacemeta-llama/Llama-3.2-1B-InstructQwen/Qwen3.8-27B

Quality

Fieldmeta-llama__llama-3.2-1b-instructqwen__qwen3.8-27b
Quality index (0–100)87.4

Cost

Fieldmeta-llama__llama-3.2-1b-instructqwen__qwen3.8-27b
Input $ / 1M tokens$0.03$0.42
Cached input $ / 1M$0.09
Output $ / 1M tokens$0.20$3.00

Context

Fieldmeta-llama__llama-3.2-1b-instructqwen__qwen3.8-27b
Max context tokens60,0001,000,000
Max output tokens54,000131,072

Modalities

Fieldmeta-llama__llama-3.2-1b-instructqwen__qwen3.8-27b
Inputtexttext, image, video
Outputtexttext