Currently listed through these providers:
Model details
Kimi K2.5
Kimi K2.5 is an open-weight multimodal model from Moonshot AI that extends the original Kimi K2 base with continued pretraining over roughly fifteen trillion mixed visual and text tokens, producing a native vision-and-language system rather than a separately bolted-on vision encoder. It is released on Hugging Face under a Modified MIT license, which gives researchers and developers room to fine-tune, inspect, and self-host the weights. The model is positioned around four intertwined capabilities: instant and thinking inference modes, conversational and agentic interaction styles, general reasoning, and visual coding, with an explicit emphasis on tool-calling and structured agent workflows that can be orchestrated through its own self-directed agent swarm paradigm.
In practical terms, Kimi K2.5 is aimed at builders who need one model that can read documents and images, reason over long contexts, and reliably drive multi-step tool use. The two hundred sixty-two thousand token context window makes it well suited to large codebases, long reports, and extended agent traces, while the visual coding strength lets it translate screenshots, diagrams, and UI mocks into working code. With open weights, long context, and a native multimodal core, it fits teams building autonomous coding assistants, document analysis pipelines, or research agents that need a single backbone for perception, planning, and action.
Quick Info
Powered by- Provider
- Ofox
- Model key
- moonshotai/kimi-k2.5
- Release date
- Jan 1, 2026
- Last updated
- Jan 1, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $3.00
Limits
- Output tokens
- 262,144 tokens
- Context window
- 262,144 tokens