Qwen2.5 72B Instruct is 5.9× cheaper per million tokens (blended). Gemini 3.7 Flash has the larger context window (1,048,576 vs 32,768 tokens, 32.0×). Qwen2.5 72B Instruct is open-weights; the other is not.

Gemini 3.7 Flash vs Qwen2.5 72B Instruct

Live catalog fields. Best-in-row is highlighted.

Gemini 3.7 Flash

Google · google__gemini-3.7-flash

Qwen2.5 72B Instructopen

Qwen · qwen__qwen-2.5-72b-instruct

Identity

Fieldgoogle__gemini-3.7-flashqwen__qwen-2.5-72b-instruct
ProviderGoogleQwen
Sluggoogle__gemini-3.7-flashqwen__qwen-2.5-72b-instruct
Statuslivelive
Open weightsNoYes
License
HuggingFaceQwen/Qwen2.5-72B-Instruct

Quality

Fieldgoogle__gemini-3.7-flashqwen__qwen-2.5-72b-instruct
Quality index (0–100)86.7

Cost

Fieldgoogle__gemini-3.7-flashqwen__qwen-2.5-72b-instruct
Input $ / 1M tokens$0.75$0.36
Cached input $ / 1M$0.07
Output $ / 1M tokens$3.75$0.40

Context

Fieldgoogle__gemini-3.7-flashqwen__qwen-2.5-72b-instruct
Max context tokens1,048,57632,768
Max output tokens65,53616,384

Modalities

Fieldgoogle__gemini-3.7-flashqwen__qwen-2.5-72b-instruct
Inputtext, image, video, file, audiotext
Outputtexttext