Qwen3.5-Flash

Qwen · qwen/qwen3.5-flash-02-23

← Back to leaderboard

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

Qwen3.5-Flash ranks #90 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.16 blended per million tokens, it is cheaper than 85% of models with published API pricing. Its 1,000,000-token context window is 3.8× the catalog median (262,144 tokens).

imagetexttext+image+video->textvideo

Context

Max context: 1000000
Max output: 65536

Pricing

Input / 1M: 0.07
Output / 1M: 0.26
Blend / 1M: 0.16

Quality

Quality index:

Provider

Provider: Qwen
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.16per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.16
10,000,000$1.63
100,000,000$16.25

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from Qwen

Closest API price