Gemma 4 31B

Google · google/gemma-4-31b-it

← Back to leaderboard

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Gemma 4 31B ranks #33 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.21 blended per million tokens, it is cheaper than 80% of models with published API pricing. Its 262,144-token context window is about the catalog median (262,144 tokens). A quality index of 76.4 places it at #33 of 59 models with an Arena Elo signal.

open weightsimagetexttext+image+video->textvideo

Context

Max context: 262144
Max output: 16384

Pricing

Input / 1M: 0.09
Output / 1M: 0.34
Blend / 1M: 0.21

Quality

Quality index: 76.4

Provider

Provider: Google
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.21per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.21
10,000,000$2.15
100,000,000$21.50

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from Google

Closest API price

Closest quality index