Gemini 3.1 Flash Lite

Google · google/gemini-3.1-flash-lite

← Back to leaderboard

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Gemini 3.1 Flash Lite ranks #190 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.88 blended per million tokens, it is cheaper than 55% of models with published API pricing. Its 1,048,576-token context window is 4.0× the catalog median (262,144 tokens).

audiofileimagetexttext+image+file+audio+video->textvideo

Context

Max context: 1048576
Max output: 65536

Pricing

Input / 1M: 0.25
Output / 1M: 1.50
Blend / 1M: 0.88

Quality

Quality index:

Provider

Provider: Google
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.88per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.88
10,000,000$8.75
100,000,000$87.50

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from Google