Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for meta-llama

Meta: Llama 4 Maverick

meta-llama/llama-4-maverick

Model weights
Compare

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward pass (400B total). It supports multilingual text and image input, and produces multilingual text and code output across 12 supported languages. Optimized for vision-language tasks, Maverick is instruction-tuned for assistant-like behavior, image reasoning, and general-purpose multimodal interaction.

Maverick features early fusion for native multimodality and a 1 million token context window. It was trained on a curated mixture of public, licensed, and Meta-platform data, covering ~22 trillion tokens, with a knowledge cutoff in August 2024. Released on April 5, 2025 under the Llama 4 Community License, Maverick is suited for research and commercial applications requiring advanced multimodal understanding and high model throughput.

Modalities

In / Out Price

$0.20 / $0.696per 1M

Context

1M

Released

Apr 5, 2025

Knowledge Cutoff

Aug 2024

Compare

About Meta: Llama 4 Maverick

OpenRouter makes Meta: Llama 4 Maverick by Meta Llama available through a unified, OpenAI-compatible API using the model ID meta-llama/llama-4-maverick. Requests can be routed across 5 providers, including DigitalOcean, DeepInfra, NovitaAI, Parasail and Google Vertex, with automatic failover when an endpoint is unavailable.

Meta: Llama 4 Maverick accepts text and images and returns text. It has a 1,048,576-token context window.

On OpenRouter, Meta: Llama 4 Maverick costs $0.20/M input tokens and $0.696/M output tokens. It was released on April 5, 2025; its knowledge cutoff is August 31, 2024.

More models from Meta Llama

  • Llama Guard 4 12B
  • Llama 4 Scout
  • Llama 3.3 70B Instruct

Frequently asked questions

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward pass (400B total). It supports multilingual text and image input, and produces multilingual text and code output across 12 supported languages.

Llama 4 Maverick costs $0.20/M input tokens and $0.696/M output tokens.

Llama 4 Maverick has a 1,048,576 token context window.

Yes. Llama 4 Maverick accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.

Llama 4 Maverick accepts text and images as input and returns text.

Llama 4 Maverick is served by 5 providers on OpenRouter: DigitalOcean, DeepInfra, NovitaAI, Parasail and Google Vertex. Requests are routed to the best available provider, with automatic failover to the others, and you can pin or exclude providers with provider routing.

Llama 4 Maverick was released on April 5, 2025. Its knowledge cutoff is August 31, 2024.