Currently listed through these providers:
Model details
Kimi K2.7 Code Highspeed
Kimi K2.7 Code Highspeed serves as the throughput-optimized variant of Moonshot's dedicated coding model. Positioned within the Kimi K2 family lineage, this high-speed version shares the same underlying architecture as the standard K2.7 Code release but prioritizes faster token generation to suit interactive development scenarios. The model functions as a specialized coding assistant rather than a general-purpose conversational system, building on instruction-following refinements that distinguish the K2.7 generation from its predecessor.
The practical advantage of this variant lies in its accelerated output speed, reaching approximately 180 tokens per second under typical conditions and up to 260 tokens per second in shorter context scenarios. The model accepts multimodal inputs including text, image, and video while producing text outputs, and operates with reasoning enabled at all times. The expansive 262,144-token context window supports long-horizon coding tasks such as multi-file refactoring and sustained debugging sessions across complex codebases. Developers working on agentic coding pipelines, rapid prototyping, or IDE-integrated workflows benefit most from the reduced latency, though resource availability may cause occasional throughput fluctuations as Moonshot scales deployment.
Quick Info
Powered by- Provider
- EmpirioLabs AI
- Model key
- kimi-k2-7-code-highspeed
- Release date
- Jun 12, 2026
- Last updated
- Jun 12, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.90
- Output token cost
- $8.00
Limits
- Output tokens
- 131,072 tokens
- Context window
- 256,000 tokens
Latest news about Kimi K2.7 Code Highspeed
Videos about Kimi K2.7 Code Highspeed
More models around Kimi K2.7 Code Highspeed
This exact model name is also listed by 9 other providers.