Qwen3 32B
Qwen · qwen/qwen3-32b
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Qwen3 32B ranks #93 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.18 blended per million tokens, it is cheaper than 84% of models with published API pricing. Its 40,960-token context window is 6.4× smaller than the catalog median (262,144 tokens).
open weightstexttext->text
Context
Max context: 40960
Max output: 16384
Max output: 16384
Pricing
Input / 1M: 0.08
Output / 1M: 0.28
Blend / 1M: 0.18
Output / 1M: 0.28
Blend / 1M: 0.18
Quality
Quality index: —
Provider
Provider: Qwen
Moderated: no
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.18per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.18 |
| 10,000,000 | $1.80 |
| 100,000,000 | $18.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Qwen
- Qwen3.8 Max (0902)$4.00 / 1M · Q 93.0
- Qwen3.8 27B$1.71 / 1M · Q 87.4
- Qwen3.7 Max$2.95 / 1M · Q 82.3
- Qwen3.6 Max Preview$3.59 / 1M · Q 79.9
- Qwen3.6 Plus$1.14 / 1M · Q 78.7
Closest API price
- Llama Guard 4 12B$0.18 / 1M
- Hy-MT2-30B-A3B$0.18 / 1M
- Hy-MT2-7B$0.18 / 1M
- Qwen3 Coder 30B A3B Instruct$0.18 / 1M
- Seed 1.6 Flash$0.19 / 1M