Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for qwen

Qwen: Qwen3 4B

qwen/qwen3-4b:free

Model weights

Qwen3-4B is a 4 billion parameter dense language model from the Qwen3 series, designed to support both general-purpose and reasoning-intensive tasks. It introduces a dual-mode architecture—thinking and non-thinking—allowing dynamic switching between high-precision logical reasoning and efficient dialogue generation. This makes it well-suited for multi-turn chat, instruction following, and complex agent workflows.

Modalities

Price

Free

Context

128K

Released

Apr 30, 2025

Knowledge Cutoff

Mar 2025

ActivityFAQExplore

Activity

Token volume and request traffic to this model over time.

Explore more models

AI Model RankingsRanking

Frequently asked questions

Qwen3-4B is a 4 billion parameter dense language model from the Qwen3 series, designed to support both general-purpose and reasoning-intensive tasks. It introduces a dual-mode architecture—thinking and non-thinking—allowing dynamic switching between high-precision logical reasoning and efficient dialogue generation.

Yes. The pricing shown on this page for Qwen3 4B is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.

Qwen3 4B has a 128,000 token context window.

Qwen3.8 27B, Qwen3.8 2.4T A95B, Qwen3.8 Max and 47 more are other text models from Qwen.

Qwen3 4B was released on April 30, 2025. Its knowledge cutoff is March 31, 2025.