Skip to content
  • Models
  • Rankings
  • Ori
Sign Up
Sign Up
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Tools
  • Pricing
  • Business
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube

Models

Models

CompareDiscover Models
Favicon for anthropic
Favicon for openai
  • Favicon for qwen
    Qwen: Qwen3.8 Max PrimeQwen3.8 Max Prime
    267M tokens

    Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video input and returns text, with a 1M-token context window and reasoning enabled by default. Tool calling, structured outputs, and configurable reasoning effort are supported, matching Qwen3.8 Max.

    by qwenSep 23, 20261M context$4/M input tokens$12/M output tokens
  • Favicon for recraft
    Recraft: Recraft V4.1 FlashRecraft V4.1 Flash
    10.4M tokens

    Recraft V4.1 Flash is a text-to-image model from Recraft, the speed and cost tier of the V4.1 family. It generates ~1K raster images in about 1.5 seconds end to end, at roughly a fifth of the cost of Recraft V4.1, and is suited for rapid iteration, prototyping, and high-volume generation where turnaround matters more than peak fidelity. It is generation only: image inputs, styles, and style references are not supported. Supports the same aspect ratios as V4.1 and the image_config parameters rgb_colors (sets a color palette) and background_rgb_color (sets the background color). See the image generation docs for details: https://openrouter.ai/docs/features/multimodal/image-generation

    by recraftSep 23, 202666K context$0.007/image
  • Favicon for stealth
    Space Bunny AlphaSpace Bunny Alpha
    148B tokens

    Space Bunny Alpha is an anonymous large model with blazing-fast inference, strong coding capabilities and native multimodal input support. It delivers adjustable reasoning effort, and a 1M-token context window. Space Bunny Alpha is a stealth model. It is developed and operated by a third-party provider who has chosen to remain anonymous during this preview. OpenRouter routes requests to it and is not its developer, owner, or provider. Prompts and completions may be retained by the provider but are not used for training; all other use is governed by the Stealth Model Terms.

    by stealthSep 23, 20261M context$0/M input tokens$0/M output tokens
  • Favicon for aion-labs
    AionLabs: Aion 3.5 MiniAion 3.5 Mini
    43.2M tokens

    Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses the same collaborative generation process in which multiple specialized models each contribute to a response, producing stronger narrative structure and more compelling tension and conflict.

    by aion-labsSep 23, 2026262K context$0.70/M input tokens$1.40/M output tokens
  • Favicon for aion-labs
    AionLabs: Aion 3.5Aion 3.5
    45.6M tokens

    Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute to a response, producing stronger narrative structure and more compelling tension and conflict.

    by aion-labsSep 23, 2026262K context$3/M input tokens$6/M output tokens
  • Favicon for upstage
    Upstage: Solar Mini 4Solar Mini 4
    50% off
    1.06B tokens

    Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context window. It is built for agentic use cases where response speed and cost matter, with fluent Korean alongside strong English and Japanese, and support for long-context workloads. When using BYOK with ZDR enforcement, the Console API key must belong to a ZDR-enabled Upstage organization. Contact Upstage to enable ZDR on your org.

    by upstageSep 23, 2026524K context$0.05/M input tokens$0.20/M output tokens
  • Favicon for cohere
    Cohere: Command A+Command A+
    8.89M tokens

    Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K context window, supports native tool calling with strict tool schemas, structured outputs (JSON object and JSON schema), optional reasoning, and configurable safety modes.

    by cohereSep 22, 2026192K context$0.30/M input tokens$1.50/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 Luna ProGPT-6 Luna Pro
    57.5B tokens

    GPT-6 Luna Pro is the same underlying model as GPT-6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

    by openaiSep 22, 20261.05M context$0.10/M input tokens$0.50/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 Luna Pro (batch)GPT-6 Luna Pro (batch)Batch variant
    995M tokens

    GPT-6 Luna Pro is the same underlying model as GPT-6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

    by openaiSep 22, 20261.05M context$0.05/M input tokens$0.25/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 LunaGPT-6 Luna
    431B tokens

    GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic tasks, and at higher reasoning effort it can take on complex software engineering and computer-use tasks that previously called for a Sol-tier model. It shares the GPT-6 family's gains in factual reliability and its clearer, more concise communication style.

    by openaiSep 22, 20261.05M context$0.10/M input tokens$0.50/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 Luna (batch)GPT-6 Luna (batch)Batch variant
    1.9B tokens

    GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic tasks, and at higher reasoning effort it can take on complex software engineering and computer-use tasks that previously called for a Sol-tier model. It shares the GPT-6 family's gains in factual reliability and its clearer, more concise communication style.

    by openaiSep 22, 20261.05M context$0.05/M input tokens$0.25/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 Sol ProGPT-6 Sol Pro
    7.06B tokens

    GPT-6 Sol Pro is the same underlying model as GPT-6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

    by openaiSep 22, 20261.05M context$2/M input tokens$10/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 Sol Pro (batch)GPT-6 Sol Pro (batch)Batch variant
    55.5M tokens

    GPT-6 Sol Pro is the same underlying model as GPT-6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

    by openaiSep 22, 20261.05M context$1/M input tokens$5/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 SolGPT-6 Sol
    137B tokens
    SEO (#34)

    GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional work, agentic coding, business workflow automation, and computer use, and is particularly strong at long-horizon software engineering tasks in real codebases. It approaches Astra-level factual reliability at a much lower cost and shares Astra's clearer, more concise communication style in technical and coding conversations.

    by openaiSep 22, 20261.05M context$2/M input tokens$10/M output tokens
  • Favicon for openai
    OpenAI: GPT-6 Sol (batch)GPT-6 Sol (batch)Batch variant
    39.2M tokens

    GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional work, agentic coding, business workflow automation, and computer use, and is particularly strong at long-horizon software engineering tasks in real codebases. It approaches Astra-level factual reliability at a much lower cost and shares Astra's clearer, more concise communication style in technical and coding conversations.

    by openaiSep 22, 20261.05M context$1/M input tokens$5/M output tokens
  • Favicon for inclusionai
    inclusionAI: Ming Image 0.1 DesignMing Image 0.1 Design
    182M tokens

    Ming Image 0.1 Design is a text-to-image model from inclusionAI aimed at graphic-design output, with an emphasis on legible text rendering inside the generated image. It generates from a prompt only and does not accept reference images. Output format can be requested as PNG, JPEG, or WebP. Image dimensions are chosen by the model rather than by the request, so explicit sizes and aspect ratios are rejected instead of silently reshaped.

    by inclusionaiSep 22, 2026$0/M input tokens$0/M output tokens
  • Favicon for anthropic
    Anthropic: Claude Opus 5.5 (batch)Claude Opus 5.5 (batch)Batch variant
    164M tokens

    Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code review and bug finding, financial and scientific analysis, and reading dense charts, diagrams, and screenshots, and it is more careful than its predecessor about only stating figures and citing sources it can back up. The model completes comparable tasks in fewer steps and with fewer tokens than Opus 5, and reports on its work in plainer language, with clear updates on what it did, what it found, and what it needs from the user. Thinking is always adaptive, so effort is the main lever for trading off depth, latency, and cost, and lower effort settings remain effective for latency-sensitive workloads.

    by anthropicSep 22, 20261M context$2/M input tokens$10/M output tokens
  • Favicon for anthropic
    Anthropic: Claude Opus 5.5Claude Opus 5.5
    200B tokens

    Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code review and bug finding, financial and scientific analysis, and reading dense charts, diagrams, and screenshots, and it is more careful than its predecessor about only stating figures and citing sources it can back up. The model completes comparable tasks in fewer steps and with fewer tokens than Opus 5, and reports on its work in plainer language, with clear updates on what it did, what it found, and what it needs from the user. Thinking is always adaptive, so effort is the main lever for trading off depth, latency, and cost, and lower effort settings remain effective for latency-sensitive workloads.

    by anthropicSep 22, 20261M context$4/M input tokens$20/M output tokens
  • Favicon for assemblyai
    AssemblyAI: Universal-3.5 ProUniversal-3.5 Pro
    50% off
    2.18M characters

    Universal-3.5 Pro is AssemblyAI's speech-to-text model served through its Sync API, returning a complete transcript with word-level timestamps in a single synchronous response for audio clips up to 120 seconds. It accepts 16-bit WAV input and supports free-text prompting, keyterms, and conversation context to steer transcription toward domain vocabulary.

    by assemblyaiSep 22, 2026from $0.000063/second
  • Favicon for xiaomi
    Xiaomi: MiMo-V2.6-Pro-UltraSpeedMiMo-V2.6-Pro-UltraSpeed
    17.6B tokens

    MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly 10x the output speed. The model features a 1M-token context window and native multimodal capabilities. Optimized for agentic workflows, it delivers top-tier performance across coding, visual, general, and research scenarios, excelling at complex, long-horizon tasks with robust generalization across a diverse range of agent harnesses.

    by xiaomiSep 21, 20261.05M context$4.35/M input tokens$8.70/M output tokens