Skip to content
  • Models
  • Rankings
  • Ori
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Tools
  • Pricing
  • Business
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status
  • AI Site Map

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube

Models

CompareDiscover Models
Favicon for anthropic
Favicon for openai

Models

CompareDiscover Models
Favicon for anthropic
Favicon for openai
  • Favicon for perplexity
    Perplexity: Decider V1.1 27BDecider V1.1 27B
    191M tokens

    Decider V1.1 27B is a new checkpoint of Perplexity's decision model, succeeding Decider V1 27B with the same API contract. Instead of generating text, it reads content passed as state (text, JSON, or images) and returns typed, probabilistic answers to one or more named questions in a single request: the probability of yes for a yes/no question (noul), a probability for every option plus the most likely one (choice), or a probability for every level of an ordered rubric plus the expected score (score). It is built for classification, routing, moderation, and rubric grading where application code thresholds the returned numbers rather than parsing a chat reply. A request can carry up to 128 questions about the same content.

    by perplexityOct 7, 2026262K context$0.02/M input tokens$0/M output tokens
  • Favicon for elevenlabs
    ElevenLabs: Eleven v4Eleven v4
    50% off
    89K tokens

    Eleven v4 is a text-to-speech model from ElevenLabs. It is ElevenLabs' most expressive model, with inline audio tags for emotional and delivery control, support for 90+ languages, and a 10,000-character request limit.

    by elevenlabsOct 7, 2026$40/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven v4 TurboEleven v4 Turbo
    50% off
    15K tokens

    Eleven v4 Turbo is a low-latency text-to-speech model from ElevenLabs. It keeps Eleven v4's expressive delivery and audio tags while being tuned for faster generation, with support for 90+ languages and a 10,000-character request limit.

    by elevenlabsOct 7, 2026$20/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven v3Eleven v3
    50% off
    18K tokens

    Eleven v3 is a text-to-speech model from ElevenLabs. It produces emotionally rich, highly expressive speech with inline audio tags, supports 70+ languages, and has a 5,000-character request limit.

    by elevenlabsOct 7, 2026$40/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven v3 ConversationalEleven v3 Conversational
    50% off
    2K tokens

    Eleven v3 Conversational is a text-to-speech model from ElevenLabs, a variant of Eleven v3 optimized for natural dialogue in conversational agents. It supports 70+ languages and has a 5,000-character request limit.

    by elevenlabsOct 7, 2026$20/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven Flash v2.5Eleven Flash v2.5
    50% off
    6K tokens

    Eleven Flash v2.5 is an ultra-low-latency text-to-speech model from ElevenLabs. It is suited for conversational and real-time use cases, supports 32 languages, and has a 40,000-character request limit.

    by elevenlabsOct 7, 2026$20/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven Turbo v2Eleven Turbo v2
    50% off
    532 tokens

    Eleven Turbo v2 is an English-only, low-latency text-to-speech model from ElevenLabs. It is suited for developer use cases where speed matters and only English is needed, and has a 30,000-character request limit.

    by elevenlabsOct 7, 2026$20/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven Multilingual v2Eleven Multilingual v2
    50% off
    6K tokens

    Eleven Multilingual v2 is a text-to-speech model from ElevenLabs. It is suited for lifelike, consistent long-form narration such as voice-overs and audiobooks, supports 29 languages, and has a 10,000-character request limit.

    by elevenlabsOct 7, 2026$40/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven Turbo v2.5Eleven Turbo v2.5
    50% off
    888 tokens

    Eleven Turbo v2.5 is a low-latency text-to-speech model from ElevenLabs. It balances quality and speed for developer use cases that need non-English languages, supports 32 languages, and has a 40,000-character request limit.

    by elevenlabsOct 7, 2026$20/M characters
  • Favicon for elevenlabs
    ElevenLabs: Eleven Flash v2Eleven Flash v2
    50% off
    557 tokens

    Eleven Flash v2 is an English-only, ultra-low-latency text-to-speech model from ElevenLabs. It is suited for conversational and real-time English use cases, and has a 30,000-character request limit.

    by elevenlabsOct 7, 2026$20/M characters
  • Favicon for elevenlabs
    ElevenLabs: Scribe v2Scribe v2
    50% off
    442K characters

    ElevenLabs Scribe v2 is a speech-to-text model that transcribes audio in 90+ languages with word-level timestamps, optional speaker diarization, and audio-event tagging.

    by elevenlabsOct 7, 2026$0.000031/second
  • Favicon for elevenlabs
    ElevenLabs: Scribe v2 MedicalScribe v2 Medical
    50% off
    10K characters

    ElevenLabs Scribe v2 Medical is a speech-to-text model tuned for clinical and medical terminology, with word-level timestamps, optional speaker diarization, and audio-event tagging.

    by elevenlabsOct 7, 2026$0.000031/second
  • Favicon for openai
    OpenAI: GPT-6 Luna DecisionsGPT-6 Luna Decisions
    7.27B tokens

    GPT-6 Luna Decisions is GPT-6 Luna served through OpenAI's Decisions API. Instead of generating text, it reads the content passed as state (text, JSON, or images) and returns typed, probabilistic answers to named questions in a single request: the probability of yes for a yes/no question (noul), a probability for every option plus the most likely one (choice), or a probability for every level of an ordered rubric plus the expected score (score). It is built for classification, routing, moderation, and rubric grading where application code thresholds the returned numbers rather than parsing a chat reply. A request can carry up to 200 questions about the same content.

    by openaiOct 6, 20261.05M context$0.10/M input tokens$0/M output tokens
  • Favicon for x-ai
    SpaceXAI: Grok Imagine Video 1.5 LiteGrok Imagine Video 1.5 Lite
    5 hours

    Grok Imagine Video 1.5 Lite is a faster, lower-cost video generation model from SpaceXAI, distilled from Grok Imagine Video 1.5. It supports text-to-video and image-to-video, trading some quality for speed and price. 1080p output is rendered at 720p and upscaled.

    by x-aiOct 6, 2026from $0.02/second
  • Favicon for google
    Google: Nano Banana 2.1Nano Banana 2.1
    316M tokens

    Nano Banana 2.1 (Gemini Nano Banana 2.1) is Google's image generation and editing model on the Flash tier, succeeding Nano Banana 2 and Nano Banana Pro. It improves product recontextualization, mask- and ink-based editing, and factual accuracy, and renders photorealistic skin tones, detailed materials, lighting, and coherent backgrounds. It accepts text and image inputs, returns images with optional text, and supports 1K, 2K, and 4K output plus extended aspect ratios via the image_config API Parameter.

    by googleOct 6, 202666K context$1.50/M input tokens$30/M output tokens
  • Favicon for mistralai
    Mistral: Mistral Large 4Mistral Large 4
    50% off
    33.3B tokens

    Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 512K-token context window with up to 256K output tokens, and supports tool calling and structured outputs.

    by mistralaiOct 6, 2026524K context$0.68/M input tokens$2.09/M output tokens
  • Favicon for tencent
    Tencent: Hy Image 3.5 PreviewHy Image 3.5 Preview
    128M tokens

    Hy Image 3.5 Preview is a unified image generation and editing model from Tencent. It handles text-to-image, image-to-image, and multi-turn editing through one endpoint, taking up to 20 reference images per request and producing output up to 4K. Built on the 80B mixture-of-experts Hy Image 3.0 base, it is particularly strong at subject consistency across edits, prompt-faithful composition, and rendering Chinese and English text inside images.

    by tencentOct 5, 2026100K context$1.60/M tokens
  • Favicon for inclusionai
    inclusionAI: Ling 3.1 FlashLing 3.1 Flash
    375B tokens

    Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.

    by inclusionaiOct 2, 2026262K context$0/M input tokens$0/M output tokens
  • Favicon for bytedance-seed
    ByteDance Seed: Seedream 5.0 FlashSeedream 5.0 Flash
    1.03B tokens

    Seedream 5.0 Flash is an image generation and editing model from ByteDance Seed. It is the fast, cost-efficient tier of the Seedream 5.0 family, suited for high-volume production and interactive editing workflows that need precise edits at low latency.

    by bytedance-seedOct 1, 2026from $0.018/image
  • Favicon for perplexity
    Perplexity: Decider V1 27BDecider V1 27B
    8.92B tokens
    Legal (#40)
    SEO (#42)
    Trivia (#33)

    Decider V1 27B is a decision model from Perplexity. Instead of generating text, it reads content passed as state and returns typed, probabilistic answers to one or more named questions in a single request: the probability of yes for a yes/no question (noul), a probability for every option plus the most likely one (choice), or a probability for every level of an ordered rubric plus the expected score (score). It is built for classification, routing, moderation, and rubric grading where application code thresholds the returned numbers rather than parsing a chat reply. A request can carry up to 128 questions about the same content. On OpenRouter it currently accepts text and JSON state; image inputs are not yet supported.

    by perplexityOct 1, 2026262K context$0.04/M input tokens$0/M output tokens