Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for z-ai

Z.ai: GLM 5.3

z-ai/glm-5.3

Compare

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves on GLM-5.2 in coding and in the balance between performance and token efficiency.

Reasoning is always on and cannot be disabled. Reasoning efforts low, high, and max are supported; max is the default.

Modalities

In / Out Price

$1.40 / $4.40per 1M

Context

1M

Released

Aug 18, 2026

Compare

About Z.ai: GLM 5.3

OpenRouter makes Z.ai: GLM 5.3 available through a unified, OpenAI-compatible API using the model ID z-ai/glm-5.3. It is served by Z.ai.

Z.ai: GLM 5.3 accepts text and returns text. It has a 1,048,576-token context window and a maximum output of 131,072 tokens.

On OpenRouter, Z.ai: GLM 5.3 costs $1.40/M input tokens and $4.40/M output tokens, with separate rates for Cache Read at $0.26/M tokens. Effective pricing can be lower when prompt caching applies. It was released on August 18, 2026.

More models from Z.ai

  • GLM 5.2
  • GLM 5.1
  • GLM 5V Turbo

Frequently asked questions

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves on GLM-5.2 in coding and in the balance between performance and token efficiency. Reasoning is always on and cannot be disabled. Reasoning efforts low, high, and max are supported; max is the default.

GLM 5.3 costs $1.40/M input tokens and $4.40/M output tokens, with separate rates for Cache Read at $0.26/M tokens.

GLM 5.3 has a 1,048,576 token context window. It supports up to 131,072 completion tokens.

Yes. GLM 5.3 accepts tools and tool_choice for function calling. It supports response_format for JSON output, without JSON-schema enforcement.

GLM 5.2, GLM 5.1, GLM 5V Turbo and 9 more are other text models from Z.ai.

GLM 5.3 was released on August 18, 2026.