Currently listed through these providers:
Model details
Kimi K2 (07/11)
Kimi K2 is a large-scale Mixture-of-Experts language model built to excel in agentic workflows, complex reasoning, and code synthesis. With a massive architecture totaling one trillion parameters, the model utilizes 32 billion active parameters per forward pass to balance computational efficiency with high-level performance. It is specifically engineered to handle demanding technical tasks, demonstrating strong capabilities across benchmarks such as LiveCodeBench for programming, SWE-bench for software engineering, and reasoning-focused evaluations like ZebraLogic and GPQA.
The model is supported by a novel training stack that incorporates the MuonClip optimizer, which ensures stability during the large-scale training of its Mixture-of-Experts architecture. Beyond its core reasoning strengths, Kimi K2 is optimized for long-context inference, allowing it to process information across a 128K token window. This design makes it a robust choice for developers and organizations requiring reliable tool-use and deep analytical processing, positioning it as a versatile tool for modern AI-driven agentic applications.
Quick Info
Powered by- Provider
- Helicone
- Model key
- kimi-k2-0711
- Release date
- Jan 1, 2025
- Last updated
- Jan 1, 2025
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.57
- Output token cost
- $2.30
Limits
- Output tokens
- 16,384 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Kimi K2 (07/11) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Kimi K2 (07/11)
No articles yet. Fetch the latest news to show it here.