GLM 5.3 Flash
Z.ai · z-ai/glm-5.3-flash
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
GLM 5.3 Flash ranks #4 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.16 blended per million tokens, it is cheaper than 85% of models with published API pricing. Its 1,048,576-token context window is 4.0× the catalog median (262,144 tokens). A quality index of 89.7 places it at #4 of 41 models with an Arena Elo signal.
Context
Max output: 131072
Pricing
Output / 1M: 0.25
Blend / 1M: 0.16
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.16per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.16 |
| 10,000,000 | $1.63 |
| 100,000,000 | $16.25 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Z.ai
- GLM 5.1$2.00 / 1M · Q 83.3
- GLM 5V Turbo$2.60 / 1M · Q 77.1
- GLM 4.7 Flash$0.23 / 1M
- GLM 4.5 Air$0.49 / 1M
- GLM 4.6V$0.60 / 1M
Closest API price
- GLM Flash Latest$0.16 / 1M
- Qwen3.5-Flash$0.16 / 1M
- Muse Spark 1.3 Contributor$0.15 / 1M
- Muse Spark 1.2 Contributor$0.15 / 1M
- Ministral 3 8B 2512$0.15 / 1M
Closest quality index
- Qwen3.8 27B$1.71 / 1M · Q 89.0
- Gemini 3.7 Flash$2.25 / 1M · Q 88.4
- Claude Fable 5$30.00 / 1M · Q 91.1
- DeepSeek V4 Pro 0423$1.12 / 1M · Q 88.2
- DeepSeek V4 Pro 0813$2.24 / 1M · Q 88.2