Search/
Skip to content
/
OpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Models
  • Providers
  • Pricing
  • Enterprise

Company

  • About
  • Announcements
  • CareersHiring
  • Privacy
  • Terms of Service
  • Support
  • State of AI
  • Works With OR

Developer

  • Documentation
  • API Reference
  • SDK
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Collections/Image Models

Image Generation Models

Model rankings updated February 2026 based on real usage data.

OpenRouter provides access to leading image generation models through a single, unified API gateway. Compare pricing, capabilities, and performance across multiple image generation APIs to find the best fit for models that transform text descriptions into high-quality images.

Image Generation Models on OpenRouter

Favicon for google

Google: Gemini 2.5 Flash Image (Nano Banana)

6.86B tokens

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation, edits, and multi-turn conversations. Aspect ratios can be controlled with the image_config API Parameter

by google33K context$0.30/M input tokens$2.50/M output tokens$30/M tokens$1/M audio tokens
Favicon for google

Google: Nano Banana Pro (Gemini 3 Pro Image Preview)

5.83B tokens

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and high-fidelity visual synthesis. The model generates context-rich graphics, from infographics and diagrams to cinematic composites, and can incorporate real-time information via Search grounding.

It offers industry-leading text rendering in images (including long passages and multilingual layouts), consistent multi-image blending, and accurate identity preservation across up to five subjects. Nano Banana Pro adds fine-grained creative controls such as localized edits, lighting and focus adjustments, camera transformations, and support for 2K/4K outputs and flexible aspect ratios. It is designed for professional-grade design, product visualization, storyboarding, and complex multi-element compositions while remaining efficient for general image creation workflows.

by google66K context$2/M input tokens$12/M output tokens$120/M tokens$2/M audio tokens
Favicon for openai

OpenAI: GPT-5 Image Mini

1.38B tokens

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by GPT-5 Mini, with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text rendering, and detailed image editing with reduced latency and cost. It excels at high-quality visual creation while maintaining strong text understanding, making it ideal for applications that require both efficient image generation and text processing at scale.

by openai400K context$2.50/M input tokens$2/M output tokens$8/M tokens
Favicon for bytedance-seed

ByteDance Seed: Seedream 4.5

561M tokens

Seedream 4.5 is the latest in-house image generation model developed by ByteDance. Compared with Seedream 4.0, it delivers comprehensive improvements, especially in editing consistency, including better preservation of subject details, lighting, and color tone. It also enhances portrait refinement and small-text rendering. The model’s multi-image composition capabilities have been significantly strengthened, and both reasoning performance and visual aesthetics continue to advance, enabling more accurate and artistically expressive image generation.

Pricing is $0.04 per output image, regardless of size.

by bytedance-seed4K context$0/M input tokens$0/M output tokens$9.581/M tokens
Favicon for openai

OpenAI: GPT-5 Image

154M tokens

GPT-5 Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following, text rendering, and detailed image editing.

by openai400K context$10/M input tokens$10/M output tokens$40/M tokens
Favicon for sourceful

Sourceful: Riverflow V2 Pro

58.3M tokens

Riverflow V2 Pro is the most powerful variant of Sourceful's Riverflow 2.0 lineup, best for top-tier control and perfect text rendering.

The Riverflow 2.0 series represents SOTA performance on image generation and editing tasks, using an integrated reasoning model to boost reliability and tackle complex challenges.

Pricing is $0.15 per 1K/2K output image and $0.33 per 4K output image.

Additional features:

  • Custom font rendering via font_inputs ($0.03/font, max 2)
  • Image enhancement via super_resolution_references ($0.20/reference, max 4)

See the image generation docs for details: https://openrouter.ai/docs/features/multimodal/image-generation

Note: Sourceful imposes a 4.5MB request size limit, therefore it is highly recommended to pass image URLs instead of Base64 data.

by sourceful8K context$0/M input tokens$0/M output tokens$35.93/M tokens
Favicon for sourceful

Sourceful: Riverflow V2 Fast

46.4M tokens

Riverflow V2 Fast is the fastest variant of Sourceful's Riverflow 2.0 lineup, best for production deployments and latency-critical workflows.

The Riverflow 2.0 series represents SOTA performance on image generation and editing tasks, using an integrated reasoning model to boost reliability and tackle complex challenges.

Pricing is $0.02 per 1K output image and $0.04 per 2K output image. Does not support 4K image output.

Additional features:

  • Custom font rendering via font_inputs ($0.03/font, max 2)
  • Image enhancement via super_resolution_references ($0.20/reference, max 4)

See the image generation docs for details: https://openrouter.ai/docs/features/multimodal/image-generation

Note: Sourceful imposes a 4.5MB request size limit, therefore it is highly recommended to pass image URLs instead of Base64 data.

by sourceful8K context$0/M input tokens$0/M output tokens$4.79/M tokens
Favicon for sourceful

Sourceful: Riverflow V2 Standard Preview

17.8M tokens

Riverflow V2 Standard Preview is the standard variant of Sourceful's Riverflow V2 preview lineup. This preview version exceeds the performance of Riverflow 1 Family and is Sourceful's first unified text-to-image and image-to-image model family.

Pricing is $0.035 per output image, regardless of size.

Sourceful imposes a 4.5MB request size limit, therefore it is highly recommended to pass image URLs instead of Base64 data.

by sourceful8K context$0/M input tokens$0/M output tokens$8.383/M tokens
Favicon for sourceful

Sourceful: Riverflow V2 Fast Preview

15.4M tokens

Riverflow V2 Fast Preview is the fastest variant of Sourceful's Riverflow V2 preview lineup. This preview version exceeds the performance of Riverflow 1 Family and is Sourceful's first unified text-to-image and image-to-image model family.

Pricing is $0.03 per output image, regardless of size.

Sourceful imposes a 4.5MB request size limit, therefore it is highly recommended to pass image URLs instead of Base64 data.

by sourceful8K context$0/M input tokens$0/M output tokens$7.186/M tokens
Favicon for sourceful

Sourceful: Riverflow V2 Max Preview

9.88M tokens

Riverflow V2 Max Preview is the most powerful variant of Sourceful's Riverflow V2 preview lineup. This preview version exceeds the performance of Riverflow 1 Family and is Sourceful's first unified text-to-image and image-to-image model family.

Pricing is $0.075 per output image, regardless of size.

Sourceful imposes a 4.5MB request size limit, therefore it is highly recommended to pass image URLs instead of Base64 data.

by sourceful8K context$0/M input tokens$0/M output tokens$17.96/M tokens
Favicon for black-forest-labs

Black Forest Labs: FLUX.2 Klein 4B

FLUX.2 [klein] 4B is the fastest and most cost-effective model in the FLUX.2 family, optimized for high-throughput use cases while maintaining excellent image quality.

Pricing is based on the output image. The first generated megapixel is charged $0.014. Each subsequent megapixel is charged $0.001.

by black-forest-labs41K context$0/M input tokens$0/M output tokens$3.418/M tokens
Favicon for black-forest-labs

Black Forest Labs: FLUX.2 Max

FLUX.2 [max] is the new top-tier image model from Black Forest Labs, pushing image quality, prompt understanding, and editing consistency to the highest level yet.

Pricing is as follows, per the docs: Input: We charge $0.03 for each megapixel on the input (i.e. reference images for editing) Output: The first generated megapixel is charged $0.07. Each subsequent megapixel is charged $0.03.

by black-forest-labs47K context$0/M input tokens$0/M output tokens$17.09/M tokens