Skip to content
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for inclusionai

Ling-3.0-flash

inclusionai/ling-3.0-flash:free

Model weights

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token.

The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.

Modalities

Price

Free

Context

262K

Released

Jul 23, 2026

ActivityFAQExplore

Activity

Token volume and request traffic to this model over time.

Explore more models

AI Model RankingsRanking

Frequently asked questions

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.

Yes. The pricing shown on this page for Ling-3.0-flash is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.

Ling-3.0-flash has a 262,144 token context window.

Ling 3.0 Flash Fin (free) is another text model from the same author.

Ling-3.0-flash was released on July 23, 2026.