Qwen3.5 397B A17B
Qwen · qwen/qwen3.5-397b-a17b
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Qwen3.5 397B A17B ranks #40 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $1.36 blended per million tokens, it is cheaper than 43% of models with published API pricing. Its 262,144-token context window is about the catalog median (262,144 tokens). A quality index of 65.3 places it at #40 of 41 models with an Arena Elo signal.
Context
Max output: 65536
Pricing
Output / 1M: 2.34
Blend / 1M: 1.36
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $1.36per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $1.36 |
| 10,000,000 | $13.65 |
| 100,000,000 | $136.50 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Qwen
- Qwen3.8 Max (0902)$4.00 / 1M · Q 93.0
- Qwen3.8 27B$1.71 / 1M · Q 87.4
- Qwen3.7 Max$2.95 / 1M · Q 82.3
- Qwen3.6 Max Preview$3.59 / 1M · Q 79.9
- Qwen3.6 Plus$1.14 / 1M · Q 78.7
Closest API price
- GLM 4.6$1.38 / 1M
- Kimi K2.5$1.35 / 1M
- Qwen3.5-122B-A10B$1.34 / 1M
- Gemini 3.5 Flash Lite$1.40 / 1M · Q 78.0
- Nova 2 Lite$1.40 / 1M
Closest quality index
- GPT-5.4 Mini$2.63 / 1M · Q 65.5
- Gemma 4 26B A4B$0.20 / 1M · Q 64.9
- GPT-5.1$5.63 / 1M · Q 75.0
- GPT-5.2 Chat$7.88 / 1M · Q 75.2
- GPT-5.2$7.88 / 1M · Q 75.2