Sulat.com
AI models
Moonshot AI logo

Model details

Kimi K2 Thinking

Kimi K2 Thinking extends Moonshot AI's K2 series into agentic, long-horizon reasoning, positioning the model as the lineup's most advanced open reasoning release to date. Independent descriptions frame it as a trillion-parameter thinking agent model designed for deep inference and tool chaining, which makes it well suited to workflows that require sustained planning rather than quick single-turn answers. The combination of an open-weights footprint with this kind of agent-oriented design signals a shift toward models that can carry out multi-step problem solving while remaining inspectable and adaptable by developers.

In practical terms, Kimi K2 Thinking fits naturally into applications such as automated research agents and complex coding assistants, where chained tool calls and extended reasoning traces matter more than raw conversational speed. Aggregator listings confirm a 262,144-token context window paired with a 262,144-token maximum output, giving the model room to plan and reflect across very long inputs and to produce sizeable generated artifacts in a single run. Per-million-token pricing on third-party routes is reported at $0.60 for input and $2.50 for output, and availability through multiple providers helps stabilize uptime for production deployments, making the model a credible option for teams that want open agentic reasoning without giving up enterprise-style reliability.

Moonshot AIkimi-k2-thinkingkimi-thinking

Quick Info

Powered by
Provider
Moonshot AI
Model key
kimi-k2-thinking
Release date
Nov 6, 2025
Last updated
Nov 6, 2025
Knowledge cutoff
2024-08
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.60
Output token cost
$2.50

Limits

Output tokens
262,144 tokens
Context window
262,144 tokens

Latest news about Kimi K2 Thinking

Moonshot AI

CoverageBenchmark

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. $0.60 per million input tokens, $2.50 per million output tokens. 262,144 token context window, maximum output of 262,144 tokens. Higher uptime with 3 providers. Includes independen

Videos about Kimi K2 Thinking

More models around Kimi K2 Thinking