Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Kimi K2 0711

Kimi K2 is built on a mixture-of-experts architecture, utilizing a system of specialized subnetworks to achieve high computational efficiency. With 32 billion active parameters out of a trillion total, the model is designed to handle complex, multi-step workflows with high reliability. It excels in long-horizon execution, enabling agents to perform autonomous tasks that span hours or even days. Beyond general reasoning, the model is engineered for coding-driven design, allowing it to transform simple prompts into functional, full-stack applications with interactive elements and database operations.

The model lineage emphasizes test-time scaling, where it reasons step-by-step and executes hundreds of sequential tool calls without human intervention. By scaling both thinking tokens and tool-calling steps, it achieves state-of-the-art performance on benchmarks like HLE and SWE-Bench Verified. This focus on autonomous, stateful execution makes it a robust choice for developers building agents that require sustained, coherent problem-solving across complex domains such as DevOps, performance optimization, and incident response.

OpenRoutermoonshotai/kimi-k2kimi-k2

Quick Info

Powered by
Provider
OpenRouter
Model key
moonshotai/kimi-k2
Release date
Jul 11, 2025
Last updated
Jul 11, 2025
Knowledge cutoff
2024-12-31
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.57
Output token cost
$2.30

Limits

Output tokens
98,304 tokens
Context window
131,072 tokens

Transparent token rates

Compare Kimi K2 0711 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Kimi K2 0711

Videos about Kimi K2 0711

More models around Kimi K2 0711