Qwen3.5-35B-A3B
Qwen · qwen/qwen3.5-35b-a3b
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
Qwen3.5-35B-A3B ranks #182 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.78 blended per million tokens, it is cheaper than 57% of models with published API pricing. Its 256,000-token context window is about the catalog median (262,144 tokens).
Context
Max output: 16384
Pricing
Output / 1M: 1.25
Blend / 1M: 0.78
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.78per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.78 |
| 10,000,000 | $7.81 |
| 100,000,000 | $78.13 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Qwen
- Qwen3.8 Max (0902)$4.00 / 1M · Q 93.0
- Qwen3.8 27B$1.71 / 1M · Q 87.4
- Qwen3.7 Max$2.95 / 1M · Q 82.3
- Qwen3.6 Max Preview$3.59 / 1M · Q 79.9
- Qwen3.6 Plus$1.14 / 1M · Q 78.7
Closest API price
- R1 Distill Llama 70B$0.80 / 1M
- Qwen3.7 Plus$0.80 / 1M · Q 77.4
- LongCat 2.0$0.75 / 1M
- MiniMax M3$0.75 / 1M
- KAT-Coder-Pro V2$0.75 / 1M