Gemma 3 4B is 6.7× cheaper per million tokens (blended). Mistral Large 3 2512 (batch) has the larger context window (262,144 vs 131,072 tokens, 2.0×). Gemma 3 4B is open-weights; the other is not.

Gemma 3 4B vs Mistral Large 3 2512 (batch)

Live catalog fields. Best-in-row is highlighted.

Gemma 3 4Bopen

Google · google__gemma-3-4b-it

Mistral Large 3 2512 (batch)

Mistral · mistralai__mistral-large-2512--batch

Identity

Fieldgoogle__gemma-3-4b-itmistralai__mistral-large-2512--batch
ProviderGoogleMistral
Sluggoogle__gemma-3-4b-itmistralai__mistral-large-2512--batch
Statuslivelive
Open weightsYesNo
License
HuggingFacegoogle/gemma-3-4b-it

Quality

Fieldgoogle__gemma-3-4b-itmistralai__mistral-large-2512--batch
Quality index (0–100)

Cost

Fieldgoogle__gemma-3-4b-itmistralai__mistral-large-2512--batch
Input $ / 1M tokens$0.05$0.25
Cached input $ / 1M$0.03
Output $ / 1M tokens$0.10$0.75

Context

Fieldgoogle__gemma-3-4b-itmistralai__mistral-large-2512--batch
Max context tokens131,072262,144
Max output tokens16,384209,715

Modalities

Fieldgoogle__gemma-3-4b-itmistralai__mistral-large-2512--batch
Inputtext, imagetext, image, file
Outputtexttext