GLM 4.7 Flash
Z.ai · z-ai/glm-4.7-flash
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
GLM 4.7 Flash ranks #107 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.23 blended per million tokens, it is cheaper than 79% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 117964
Pricing
Output / 1M: 0.40
Blend / 1M: 0.23
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.23per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.23 |
| 10,000,000 | $2.30 |
| 100,000,000 | $23.02 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Z.ai
- GLM 5.3 Flash$0.16 / 1M · Q 87.9
- GLM 5.1$2.00 / 1M · Q 81.9
- GLM 5V Turbo$2.60 / 1M · Q 76.0
- GLM 4.5 Air$0.49 / 1M
- GLM 4.6V$0.60 / 1M
Closest API price
- GPT-5 Nano$0.22 / 1M
- Nemotron 3 Super$0.24 / 1M
- Gemma 4 31B$0.21 / 1M · Q 76.4
- Seed-2.0-Mini$0.25 / 1M
- Gemini 2.5 Flash Lite$0.25 / 1M