Gemini 3.5 Flash
Google · google/gemini-3.5-flash
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Gemini 3.5 Flash ranks #19 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $5.25 blended per million tokens, it is cheaper than 20% of models with published API pricing. Its 1,048,576-token context window is 4.0× the catalog median (262,144 tokens). A quality index of 81.3 places it at #19 of 41 models with an Arena Elo signal.
Context
Max output: 65536
Pricing
Output / 1M: 9.00
Blend / 1M: 5.25
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $5.25per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $5.25 |
| 10,000,000 | $52.50 |
| 100,000,000 | $525.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Google
- Gemini 3.7 Flash$2.25 / 1M · Q 86.7
- Gemini 3.6 Flash$2.25 / 1M · Q 83.7
- Gemini 3.8 Flash$2.25 / 1M · Q 80.8
- Gemini 3.1 Pro Preview$7.00 / 1M · Q 80.4
- Gemini 3.5 Flash Lite$1.40 / 1M · Q 78.0
Closest API price
- o3$5.00 / 1M
- GPT-4.1$5.00 / 1M
- Sonar Reasoning Pro$5.00 / 1M
- Sonar Deep Research$5.00 / 1M
- Gemini 2.5 Pro$5.63 / 1M · Q 76.2
Closest quality index
- Claude Opus 4.5$15.00 / 1M · Q 80.9
- Gemini 3.8 Flash$2.25 / 1M · Q 80.8
- GLM 5.1$2.00 / 1M · Q 81.9
- Gemini 3.1 Pro Preview$7.00 / 1M · Q 80.4
- GPT-5.5$17.50 / 1M · Q 80.4