Currently listed through these providers:
Model details
Kimi K2.7 Code
Kimi K2.7 Code is positioned as a specialized coding model in the Kimi K2 family, designed to improve instruction compliance and long-horizon programming performance over the previous K2.6 release. According to the Kimi API documentation, external benchmark evaluations report meaningful gains on instruction following for coding tasks and a roughly 30% average reduction in overthinking tendencies, signaling a focus on producing more decisive, task-oriented outputs rather than verbose deliberation. The model is also distributed as open weights, which makes it usable for self-hosting, fine-tuning, or integration into developer tooling that benefits from transparent model artifacts.
A practical highlight of the release is a high-speed sibling variant, kimi-k2.7-code-highspeed, which shares the same underlying model but is optimized for throughput, delivering around 180 tokens per second and up to roughly 260 tokens per second in short-context scenarios. That makes it well suited to interactive coding assistants where latency matters, such as code completion, refactoring chat, and iterative debugging inside an IDE. The model has also reached broader distribution through GitHub Copilot, where it is offered as the first open-weight option in the model picker, hosted on Microsoft Azure and billed at provider list pricing, extending its reach to teams that prefer open-weight choices in their coding workflows.
Quick Info
Powered by- Provider
- Neuralwatt
- Model key
- kimi-k2.7-code
- Release date
- Jun 12, 2026
- Last updated
- Jun 12, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.95
- Output token cost
- $4.00
Limits
- Output tokens
- 262,128 tokens
- Context window
- 262,128 tokens
Latest news about Kimi K2.7 Code
Videos about Kimi K2.7 Code
More models around Kimi K2.7 Code
This exact model name is also listed by 59 other providers.