Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for z-ai

Z.ai: GLM 5 Turbo

z-ai/glm-5-turbo

Compare

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows involving long execution chains, with improved complex instruction decomposition, tool use, scheduled and persistent execution, and overall stability across extended tasks.

Modalities

In / Out Price

$1.20 / $4per 1M

Context

203K

Released

Mar 15, 2026

Compare

About Z.ai: GLM 5 Turbo

OpenRouter makes Z.ai: GLM 5 Turbo available through a unified, OpenAI-compatible API using the model ID z-ai/glm-5-turbo. It is served by Z.ai.

Z.ai: GLM 5 Turbo accepts text and returns text. It has a 202,752-token context window and a maximum output of 131,072 tokens.

On OpenRouter, Z.ai: GLM 5 Turbo costs $1.20/M input tokens and $4.00/M output tokens, with separate rates for Cache Read at $0.24/M tokens. Effective pricing can be lower when prompt caching applies. It was released on March 15, 2026.

More models from Z.ai

  • GLM 5.3
  • GLM 5.2
  • GLM 5.1

Frequently asked questions

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows involving long execution chains, with improved complex instruction decomposition, tool use, scheduled and persistent execution, and overall stability across extended tasks.

GLM 5 Turbo costs $1.20/M input tokens and $4.00/M output tokens, with separate rates for Cache Read at $0.24/M tokens.

GLM 5 Turbo has a 202,752 token context window. It supports up to 131,072 completion tokens.

Yes. GLM 5 Turbo accepts tools and tool_choice for function calling. It supports response_format for JSON output, without JSON-schema enforcement.

GLM 5.3, GLM 5.2, GLM 5.1 and 9 more are other text models from Z.ai.

GLM 5 Turbo was released on March 15, 2026.