
Qwen3-0.6B is a lightweight, 0.6 billion parameter language model in the Qwen3 series, offering support for both general-purpose dialogue and structured reasoning through a dual-mode (thinking/non-thinking) architecture. Despite its small size, it supports long contexts up to 32,768 tokens and provides multilingual, tool-use, and instruction-following capabilities.
Modalities
Price
Free
Context
32K
Released
Apr 30, 2025
Knowledge Cutoff
Mar 2025
Token volume and request traffic to this model over time.
Qwen3-0.6B is a lightweight, 0.6 billion parameter language model in the Qwen3 series, offering support for both general-purpose dialogue and structured reasoning through a dual-mode (thinking/non-thinking) architecture. Despite its small size, it supports long contexts up to 32,768 tokens and provides multilingual, tool-use, and instruction-following capabilities.
Yes. The pricing shown on this page for Qwen3 0.6B is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
Qwen3 0.6B has a 32,000 token context window.
Qwen3.8 Max Prime, Qwen3.8 Omni Flash, Qwen3.8 Max (0902) and 50 more are other text models from Qwen.
Qwen3 0.6B was released on April 30, 2025. Its knowledge cutoff is March 31, 2025.