
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for tasks like math, programming, and logical inference, and a "non-thinking" mode for general-purpose conversation. The model is fine-tuned for instruction-following, agent tool use, creative writing, and multilingual tasks across 100+ languages and dialects. It natively handles 32K token contexts and can extend to 131K tokens using YaRN-based scaling.
Modalities
Price
Free
Context
132K
Released
Apr 28, 2025
Knowledge Cutoff
Mar 2025
Token volume and request traffic to this model over time.
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for tasks like math, programming, and logical inference, and a "non-thinking" mode for general-purpose conversation.
Yes. The pricing shown on this page for Qwen3 14B is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
Qwen3 14B has a 131,702 token context window.
Qwen3.8 Flash, Qwen3.8 27B, Qwen3.8 2.4T A95B and 47 more are other text models from Qwen.
Qwen3 14B was released on April 28, 2025. Its knowledge cutoff is March 31, 2025.