Gemini 2.5 Flash Lite

Google · google/gemini-2.5-flash-lite

← Back to leaderboard

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Gemini 2.5 Flash Lite ranks #76 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.25 blended per million tokens, it is cheaper than 78% of models with published API pricing. Its 1,048,576-token context window is 4.0× the catalog median (262,144 tokens).

audiofileimagetexttext+image+file+audio+video->textvideo

Context

Max context: 1048576
Max output: 65535

Pricing

Input / 1M: 0.10
Output / 1M: 0.40
Blend / 1M: 0.25

Quality

Quality index:

Provider

Provider: Google
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.25per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.25
10,000,000$2.50
100,000,000$25.00

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from Google

Closest API price