Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for moonshotai

MoonshotAI: Kimi K2 Thinking

moonshotai/kimi-k2-thinking

Model weights
Compare

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in Kimi K2, it activates 32 billion parameters per forward pass and supports 256 k-token context windows. The model is optimized for persistent step-by-step thought, dynamic tool invocation, and complex reasoning workflows that span hundreds of turns. It interleaves step-by-step reasoning with tool use, enabling autonomous research, coding, and writing that can persist for hundreds of sequential actions without drift.

It sets new open-source benchmarks on HLE, BrowseComp, SWE-Multilingual, and LiveCodeBench, while maintaining stable multi-agent behavior through 200–300 tool calls. Built on a large-scale MoE architecture with MuonClip optimization, it combines strong reasoning depth with high inference efficiency for demanding agentic and analytical tasks.

Modalities

In / Out Price

$0.60 / $2.50per 1M

Context

262K

Released

Nov 6, 2025

Compare

About MoonshotAI: Kimi K2 Thinking

OpenRouter makes MoonshotAI: Kimi K2 Thinking available through a unified, OpenAI-compatible API using the model ID moonshotai/kimi-k2-thinking. Requests can be routed across 2 providers, including NovitaAI and Google Vertex, with automatic failover when an endpoint is unavailable.

MoonshotAI: Kimi K2 Thinking accepts text and returns text. It has a 262,144-token context window and a maximum output of 100,352 tokens.

On OpenRouter, MoonshotAI: Kimi K2 Thinking costs $0.60/M input tokens and $2.50/M output tokens, with separate rates for Cache Read at $0.15/M tokens. Effective pricing can be lower when prompt caching applies. It was released on November 6, 2025.

More models from MoonshotAI

  • Kimi K3
  • Kimi K2.7 Code
  • Kimi K2.6

Frequently asked questions

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in Kimi K2, it activates 32 billion parameters per forward pass and supports 256 k-token context windows.

Kimi K2 Thinking costs $0.60/M input tokens and $2.50/M output tokens, with separate rates for Cache Read at $0.15/M tokens.

Kimi K2 Thinking has a 262,144 token context window. It supports up to 100,352 completion tokens.

Yes. Kimi K2 Thinking accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.

Kimi K2 Thinking is served by 2 providers on OpenRouter: NovitaAI and Google Vertex. Requests are routed to the best available provider, with automatic failover to the others, and you can pin or exclude providers with provider routing.

Kimi K2 Thinking was released on November 6, 2025.