Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support across 140+ languages. Strong on coding, reasoning, and document understanding tasks. Apache 2.0 license.
Modalities
In / Out Price
$0.08 / $0.35per 1M
Context
262K
Released
Apr 2, 2026
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support across 140+ languages. Strong on coding, reasoning, and document understanding tasks. Apache 2.0 license.
Gemma 4 31B costs $0.08/M input tokens and $0.35/M output tokens, with separate rates for Cache Read at $0.01/M tokens.
Gemma 4 31B has a 262,144 token context window. It supports up to 8,192 completion tokens.
Yes. Gemma 4 31B accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.
Gemma 4 31B accepts images, text and video as input and returns text.
Gemma 4 31B is served by 15 providers on OpenRouter: OpenInference, DeepInfra Ultra, CoreWeave, Venice, Chutes, SiliconFlow, NovitaAI, Friendli and 7 more. Requests are routed to the best available provider, with automatic failover to the others, and you can pin or exclude providers with provider routing.
Gemma 4 31B was released on April 2, 2026.