Qwen3 VL 32B Instruct
Qwen · qwen/qwen3-vl-32b-instruct
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
Qwen3 VL 32B Instruct ranks #112 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.26 blended per million tokens, it is cheaper than 78% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 32768
Pricing
Output / 1M: 0.42
Blend / 1M: 0.26
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.26per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.26 |
| 10,000,000 | $2.60 |
| 100,000,000 | $26.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Qwen
- Qwen3.8 Max (0902)$4.00 / 1M · Q 93.0
- Qwen3.8 27B$1.71 / 1M · Q 87.4
- Qwen3.7 Max$2.95 / 1M · Q 82.3
- Qwen3.6 Max Preview$3.59 / 1M · Q 79.9
- Qwen3.6 Plus$1.14 / 1M · Q 78.7
Closest API price
- Gemma 3 27B$0.26 / 1M
- Hermes 4 70B$0.27 / 1M
- Seed-2.0-Mini$0.25 / 1M
- Gemini 2.5 Flash Lite$0.25 / 1M
- GPT-4.1 Nano$0.25 / 1M