Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for google

Google: Gemma 4 31B

google/gemma-4-31b-it

Model weights
Compare

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support across 140+ languages. Strong on coding, reasoning, and document understanding tasks. Apache 2.0 license.

Modalities

In / Out Price

$0.08 / $0.35per 1M

Context

262K

Released

Apr 2, 2026

Compare

About Google: Gemma 4 31B

OpenRouter makes Google: Gemma 4 31B available through a unified, OpenAI-compatible API using the model ID google/gemma-4-31b-it. Requests can be routed across 15 providers, including OpenInference, DeepInfra Ultra, CoreWeave, Venice, Chutes, SiliconFlow, NovitaAI, Friendli and 7 more, with automatic failover when an endpoint is unavailable.

Google: Gemma 4 31B accepts images, text and video and returns text. It has a 262,144-token context window and a maximum output of 8,192 tokens.

On OpenRouter, Google: Gemma 4 31B costs $0.08/M input tokens and $0.35/M output tokens, with separate rates for Cache Read at $0.01/M tokens. Effective pricing can be lower when prompt caching applies. It was released on April 2, 2026.

More models from Google

  • Gemini 3.7 Flash
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite

Frequently asked questions

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support across 140+ languages. Strong on coding, reasoning, and document understanding tasks. Apache 2.0 license.

Gemma 4 31B costs $0.08/M input tokens and $0.35/M output tokens, with separate rates for Cache Read at $0.01/M tokens.

Gemma 4 31B has a 262,144 token context window. It supports up to 8,192 completion tokens.

Yes. Gemma 4 31B accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.

Gemma 4 31B accepts images, text and video as input and returns text.

Gemma 4 31B is served by 15 providers on OpenRouter: OpenInference, DeepInfra Ultra, CoreWeave, Venice, Chutes, SiliconFlow, NovitaAI, Friendli and 7 more. Requests are routed to the best available provider, with automatic failover to the others, and you can pin or exclude providers with provider routing.

Gemma 4 31B was released on April 2, 2026.