GLM 4.7 Flash

Z.ai · z-ai/glm-4.7-flash

← Back to leaderboard

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

GLM 4.7 Flash ranks #107 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.23 blended per million tokens, it is cheaper than 79% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).

open weightstexttext->text

Context

Max context: 131072
Max output: 117964

Pricing

Input / 1M: 0.06
Output / 1M: 0.40
Blend / 1M: 0.23

Quality

Quality index:

Provider

Provider: Z.ai
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.23per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.23
10,000,000$2.30
100,000,000$23.02

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from Z.ai

Closest API price