Currently listed through these providers:
Model details
Kimi K2.7 Code Fast
Kimi K2.7 Code Fast sits inside the Kimi K2 family as a coding-oriented variant built to balance reasoning depth with responsive generation. Its design centers on developer workflows that demand step-by-step logical handling alongside fast conversational turnaround, making it a practical fit for code review, refactoring assistance, debug explanation, and structured automation tasks. The model accepts both text and image inputs while producing text outputs, which allows it to interpret screenshots, diagrams, or UI snippets as part of programming and documentation work, rather than treating every interaction as pure source code.
In real deployments Kimi K2.7 Code Fast is offered as an open-weights option, giving teams the flexibility to host and fine-tune under their own infrastructure rather than relying solely on a hosted endpoint. A listing surfaced in the Fireworks AI router catalog identifies the model under the identifier accounts/fireworks/routers/kimi-k2p7-code-fast, with a 262,000-token context window and matching output ceiling that supports long-file reasoning across large repositories, multi-turn refactors, and extended agent loops. The same listing highlights tool calling, temperature control, and reasoning capabilities, which together make the model well suited to autonomous coding pipelines that need structured outputs, controlled sampling, and the ability to invoke external functions during a single session.
Quick Info
Powered by- Provider
- Neuralwatt
- Model key
- kimi-k2.7-code-fast
- Release date
- Jun 12, 2026
- Last updated
- Jun 12, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.95
- Output token cost
- $4.00
Limits
- Output tokens
- 262,128 tokens
- Context window
- 262,128 tokens