Qwen3 VL 8B Thinking is 26.3× cheaper per million tokens (blended). Claude Fable 5 has the larger context window (1,000,000 vs 131,072 tokens, 7.6×). Qwen3 VL 8B Thinking is open-weights; the other is not.

Claude Fable 5 vs Qwen3 VL 8B Thinking

Live catalog fields. Best-in-row is highlighted.

Claude Fable 5

Anthropic · anthropic__claude-fable-5

Qwen3 VL 8B Thinkingopen

Qwen · qwen__qwen3-vl-8b-thinking

Identity

Fieldanthropic__claude-fable-5qwen__qwen3-vl-8b-thinking
ProviderAnthropicQwen
Sluganthropic__claude-fable-5qwen__qwen3-vl-8b-thinking
Statuslivelive
Open weightsNoYes
License
HuggingFaceQwen/Qwen3-VL-8B-Thinking

Quality

Fieldanthropic__claude-fable-5qwen__qwen3-vl-8b-thinking
Quality index (0–100)91.1

Cost

Fieldanthropic__claude-fable-5qwen__qwen3-vl-8b-thinking
Input $ / 1M tokens$10.00$0.18
Cached input $ / 1M$1.00
Output $ / 1M tokens$50.00$2.10

Context

Fieldanthropic__claude-fable-5qwen__qwen3-vl-8b-thinking
Max context tokens1,000,000131,072
Max output tokens128,00032,768

Modalities

Fieldanthropic__claude-fable-5qwen__qwen3-vl-8b-thinking
Inputtext, image, fileimage, text
Outputtexttext