GLM-4-32B-0414 is a 32B bilingual (Chinese-English) open-weight language model optimized for code generation, function calling, and agent-style tasks. Pretrained on 15T of high-quality and reasoning-heavy data, it was further refined using human preference alignment, rejection sampling, and reinforcement learning. The model excels in complex reasoning, artifact generation, and structured output tasks, achieving performance comparable to GPT-4o and DeepSeek-V3-0324 across several benchmarks.
Modalities
Price
Free
Context
33K
Released
Apr 17, 2025
Knowledge Cutoff
Jun 2024
Token volume and request traffic to this model over time.
GLM-4-32B-0414 is a 32B bilingual (Chinese-English) open-weight language model optimized for code generation, function calling, and agent-style tasks. Pretrained on 15T of high-quality and reasoning-heavy data, it was further refined using human preference alignment, rejection sampling, and reinforcement learning.
Yes. The pricing shown on this page for GLM 4 32B is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
GLM 4 32B has a 32,768 token context window.
GLM 4 32B was released on April 17, 2025. Its knowledge cutoff is June 30, 2024.