Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable thinking and instant modes.
Modalities
Price
Free
Context
262K
Released
Aug 6, 2026
Token volume and request traffic to this model over time.
Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable thinking and instant modes.
Yes. The pricing shown on this page for Ling 3.0 Tiny is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
Ling 3.0 Tiny has a 262,144 token context window.
Ling-3.0-flash is another text model from inclusionAI.
Ling 3.0 Tiny was released on August 6, 2026.