Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Helicone logo

Model details

Kimi K2 (07/11)

Kimi K2 is a large-scale Mixture-of-Experts language model built to excel in agentic workflows, complex reasoning, and code synthesis. With a massive architecture totaling one trillion parameters, the model utilizes 32 billion active parameters per forward pass to balance computational efficiency with high-level performance. It is specifically engineered to handle demanding technical tasks, demonstrating strong capabilities across benchmarks such as LiveCodeBench for programming, SWE-bench for software engineering, and reasoning-focused evaluations like ZebraLogic and GPQA.

The model is supported by a novel training stack that incorporates the MuonClip optimizer, which ensures stability during the large-scale training of its Mixture-of-Experts architecture. Beyond its core reasoning strengths, Kimi K2 is optimized for long-context inference, allowing it to process information across a 128K token window. This design makes it a robust choice for developers and organizations requiring reliable tool-use and deep analytical processing, positioning it as a versatile tool for modern AI-driven agentic applications.

Heliconekimi-k2-0711kimi-k2

Quick Info

Powered by
Provider
Helicone
Model key
kimi-k2-0711
Release date
Jan 1, 2025
Last updated
Jan 1, 2025
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.57
Output token cost
$2.30

Limits

Output tokens
16,384 tokens
Context window
131,072 tokens

Transparent token rates

Compare Kimi K2 (07/11) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Kimi K2 (07/11)

No articles yet. Fetch the latest news to show it here.

Videos about Kimi K2 (07/11)

More models around Kimi K2 (07/11)