As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, and tool collaboration, and has achieved leading performance among open-source models of the same size on several current public benchmark leaderboards.
Modalities
Price
Free
Context
200K
Released
Jan 19, 2026
Token volume and request traffic to this model over time.
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, and tool collaboration, and has achieved leading performance among open-source models of the same size on several current public benchmark leaderboards.
Yes. The pricing shown on this page for GLM 4.7 Flash is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
GLM 4.7 Flash has a 200,000 token context window.
GLM 5.3 Flash, GLM 5.3, GLM 5.2 and 10 more are other text models from Z.ai.
GLM 4.7 Flash was released on January 19, 2026.