Sulat.com
AI models
Ofox logo

Model details

Kimi K2.7 Code

Kimi K2.7 Code is Kimi's dedicated coding model, positioned as a focused upgrade for software-generation workloads rather than a general-purpose conversational system. The official Kimi API Platform documentation describes it as following instructions more reliably in long contexts and completing coding tasks with higher success rates, with external benchmark evaluations showing significantly improved instruction compliance and long-horizon coding performance compared to the prior K2.6 release. The same documentation reports an average reduction in overthinking tendencies of roughly 30%, a qualitative signal that the model reasons more efficiently on multi-step programming tasks rather than producing unnecessarily verbose chains of thought.

For latency-sensitive workflows, Kimi publishes a parallel high-speed variant, kimi-k2.7-code-highspeed, which is the same underlying model but tuned for faster output, delivering approximately 180 tokens per second and up to around 260 tokens per second in short-context scenarios, according to Kimi's quickstart documentation. The platform notes that capacity for this high-speed tier is currently limited and that throughput may fluctuate while resources are gradually scaled up, so teams targeting consistent generation speed should plan around that variability. Together, the standard and high-speed variants give developers a choice between thorough long-horizon coding behavior and a throughput-optimized mode for interactive development loops.

Ofoxmoonshotai/kimi-k2.7-codekimi-k2

Quick Info

Powered by
Provider
Ofox
Model key
moonshotai/kimi-k2.7-code
Release date
Jun 12, 2026
Last updated
Jun 12, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.95
Output token cost
$4.00

Limits

Output tokens
262,144 tokens
Context window
262,144 tokens

Latest news about Kimi K2.7 Code

Videos about Kimi K2.7 Code

More models around Kimi K2.7 Code