Currently listed through these providers:
Model details
Kimi K2 0711
Kimi K2 is built on a mixture-of-experts architecture, utilizing a system of specialized subnetworks to achieve high computational efficiency. With 32 billion active parameters out of a trillion total, the model is designed to handle complex, multi-step workflows with high reliability. It excels in long-horizon execution, enabling agents to perform autonomous tasks that span hours or even days. Beyond general reasoning, the model is engineered for coding-driven design, allowing it to transform simple prompts into functional, full-stack applications with interactive elements and database operations.
The model lineage emphasizes test-time scaling, where it reasons step-by-step and executes hundreds of sequential tool calls without human intervention. By scaling both thinking tokens and tool-calling steps, it achieves state-of-the-art performance on benchmarks like HLE and SWE-Bench Verified. This focus on autonomous, stateful execution makes it a robust choice for developers building agents that require sustained, coherent problem-solving across complex domains such as DevOps, performance optimization, and incident response.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- moonshotai/kimi-k2
- Release date
- Jul 11, 2025
- Last updated
- Jul 11, 2025
- Knowledge cutoff
- 2024-12-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.57
- Output token cost
- $2.30
Limits
- Output tokens
- 98,304 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Kimi K2 0711 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.