
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following, logical reasoning, math, code, and tool usage. The model supports a native 262K context length and does not implement "thinking mode" (<think> blocks).
Compared to its base variant, this version delivers significant gains in knowledge coverage, long-context reasoning, coding benchmarks, and alignment with open-ended tasks. It is particularly strong on multilingual understanding, math reasoning (e.g., AIME, HMMT), and alignment evaluations like Arena-Hard and WritingBench.
Modalities
Price
Free
Context
262K
Released
Jul 21, 2025
Knowledge Cutoff
Jun 2025
Token volume and request traffic to this model over time.
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following, logical reasoning, math, code, and tool usage.
Yes. The pricing shown on this page for Qwen3 235B A22B Instruct 2507 is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
Qwen3 235B A22B Instruct 2507 has a 262,144 token context window.
Qwen3.8 Flash, Qwen3.8 27B, Qwen3.8 2.4T A95B and 47 more are other text models from Qwen.
Qwen3 235B A22B Instruct 2507 was released on July 21, 2025. Its knowledge cutoff is June 30, 2025.