Gemma 3 4B is 10.8× cheaper per million tokens (blended). GLM 5.3 FlashX has the larger context window (1,048,576 vs 131,072 tokens, 8.0×). Gemma 3 4B is open-weights; the other is not.
Gemma 3 4B vs GLM 5.3 FlashX
Live catalog fields. Best-in-row is highlighted.
Gemma 3 4Bopen
Google · google__gemma-3-4b-it
GLM 5.3 FlashX
Z.ai · z-ai__glm-5.3-flashx
Identity
| Field | google__gemma-3-4b-it | z-ai__glm-5.3-flashx |
|---|---|---|
| Provider | Z.ai | |
| Slug | google__gemma-3-4b-it | z-ai__glm-5.3-flashx |
| Status | live | live |
| Open weights | Yes | No |
| License | — | — |
| HuggingFace | google/gemma-3-4b-it | — |
Quality
| Field | google__gemma-3-4b-it | z-ai__glm-5.3-flashx |
|---|---|---|
| Quality index (0–100) | — | — |
Cost
| Field | google__gemma-3-4b-it | z-ai__glm-5.3-flashx |
|---|---|---|
| Input $ / 1M tokens | $0.05 | $0.37 |
| Cached input $ / 1M | — | $0.07 |
| Output $ / 1M tokens | $0.10 | $1.25 |
Context
| Field | google__gemma-3-4b-it | z-ai__glm-5.3-flashx |
|---|---|---|
| Max context tokens | 131,072 | 1,048,576 |
| Max output tokens | 16,384 | 131,072 |
Modalities
| Field | google__gemma-3-4b-it | z-ai__glm-5.3-flashx |
|---|---|---|
| Input | text, image | text, image, video |
| Output | text | text |