Qwen3 32B

Qwen · qwen/qwen3-32b

← Back to leaderboard

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

Qwen3 32B ranks #93 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.18 blended per million tokens, it is cheaper than 84% of models with published API pricing. Its 40,960-token context window is 6.4× smaller than the catalog median (262,144 tokens).

open weightstexttext->text

Context

Max context: 40960
Max output: 16384

Pricing

Input / 1M: 0.08
Output / 1M: 0.28
Blend / 1M: 0.18

Quality

Quality index:

Provider

Provider: Qwen
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.18per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.18
10,000,000$1.80
100,000,000$18.00

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from Qwen

Closest API price