Qwen3 30B A3B Instruct 2507
Qwen · qwen/qwen3-30b-a3b-instruct-2507
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
Qwen3 30B A3B Instruct 2507 ranks #75 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.12 blended per million tokens, it is cheaper than 90% of models with published API pricing. Its 128,000-token context window is 2.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 32000
Pricing
Output / 1M: 0.19
Blend / 1M: 0.12
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.12per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.12 |
| 10,000,000 | $1.21 |
| 100,000,000 | $12.06 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Qwen
- Qwen3.8 Max (0902)$4.00 / 1M · Q 93.0
- Qwen3.8 27B$1.71 / 1M · Q 87.4
- Qwen3.7 Max$2.95 / 1M · Q 82.3
- Qwen3.6 Max Preview$3.59 / 1M · Q 79.9
- Qwen3.6 Plus$1.14 / 1M · Q 78.7
Closest API price
- Ling 3.0 Flash Fin$0.12 / 1M
- DeepSeek V4 Flash 0731$0.12 / 1M · Q 86.3
- Nemotron 3 Nano 30B A3B$0.12 / 1M
- Granite 4.2 8B$0.13 / 1M
- Qwen3.5-9B$0.13 / 1M