GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows involving long execution chains, with improved complex instruction decomposition, tool use, scheduled and persistent execution, and overall stability across extended tasks.
Modalities
In / Out Price
$1.20 / $4per 1M
Context
203K
Released
Mar 15, 2026
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows involving long execution chains, with improved complex instruction decomposition, tool use, scheduled and persistent execution, and overall stability across extended tasks.
GLM 5 Turbo costs $1.20/M input tokens and $4.00/M output tokens, with separate rates for Cache Read at $0.24/M tokens.
GLM 5 Turbo has a 202,752 token context window. It supports up to 131,072 completion tokens.
Yes. GLM 5 Turbo accepts tools and tool_choice for function calling. It supports response_format for JSON output, without JSON-schema enforcement.
GLM 5 Turbo was released on March 15, 2026.