Gemini 2.5 Flash Lite
Google · google/gemini-2.5-flash-lite
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Gemini 2.5 Flash Lite ranks #76 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.25 blended per million tokens, it is cheaper than 78% of models with published API pricing. Its 1,048,576-token context window is 4.0× the catalog median (262,144 tokens).
Context
Max output: 65535
Pricing
Output / 1M: 0.40
Blend / 1M: 0.25
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.25per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.25 |
| 10,000,000 | $2.50 |
| 100,000,000 | $25.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Google
- Lyria 3 Pro Preview$0.00 / 1M
- Lyria 3 Clip Preview$0.00 / 1M
- Gemma 3 4B$0.07 / 1M
- Gemma 3 12B$0.10 / 1M
- Gemma 4 26B A4B$0.20 / 1M
Closest API price
- Seed-2.0-Mini$0.25 / 1M
- GPT-4.1 Nano$0.25 / 1M
- Nemotron 3 Super$0.24 / 1M
- Qwen3 VL 32B Instruct$0.26 / 1M
- Gemma 3 27B$0.26 / 1M