GLM 4.5
Z.ai · z-ai/glm-4.5
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
GLM 4.5 ranks #233 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $1.40 blended per million tokens, it is cheaper than 41% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 98304
Pricing
Output / 1M: 2.20
Blend / 1M: 1.40
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $1.40per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $1.40 |
| 10,000,000 | $14.00 |
| 100,000,000 | $140.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Z.ai
- GLM 5.3 Flash$0.16 / 1M · Q 89.7
- GLM 5.1$2.00 / 1M · Q 83.3
- GLM 5V Turbo$2.60 / 1M · Q 77.1
- GLM 4.7 Flash$0.23 / 1M
- GLM 4.5 Air$0.49 / 1M
Closest API price
- Gemini 3.5 Flash Lite$1.40 / 1M · Q 79.3
- Nova 2 Lite$1.40 / 1M
- Nano Banana (Gemini 2.5 Flash Image)$1.40 / 1M
- Morph V3 Large$1.40 / 1M
- Gemini 2.5 Flash$1.40 / 1M