Qwen3.8 Omni Flash
Qwen · qwen/qwen3.8-omni-flash
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
Qwen3.8 Omni Flash ranks #125 of 374 models in the default AI Benchmark Hub order (quality first, then cost). At $0.31 blended per million tokens, it is cheaper than 75% of models with published API pricing. Its 1,000,000-token context window is 3.8× the catalog median (262,144 tokens).
Context
Max output: 131072
Pricing
Output / 1M: 0.47
Blend / 1M: 0.31
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.31per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.31 |
| 10,000,000 | $3.10 |
| 100,000,000 | $31.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Qwen
- Qwen3.8 Max (0902)$4.00 / 1M · Q 92.9
- Qwen3.8 27B$1.71 / 1M · Q 87.2
- Qwen3.7 Max$2.95 / 1M · Q 82.2
- Qwen3.6 Max Preview$3.59 / 1M · Q 79.8
- Qwen3.7 Plus$0.80 / 1M · Q 77.6
Closest API price
- Qwen3.8 Flash$0.31 / 1M
- Qwen3 30B A3B$0.31 / 1M
- GPT-6 Luna Pro$0.30 / 1M
- GPT-6 Luna$0.30 / 1M
- GPT Luna Latest$0.30 / 1M