Sulat.com
AI models
Get 10-25% off from Qwen
Alibaba Coding Plan logo

Model details

Kimi K2.5

Kimi K2.5 is an open-source native multimodal agentic model designed to unify vision and language understanding within a single architecture. Built through continual pretraining on approximately 15 trillion mixed visual and text tokens atop the Kimi-K2-Base foundation, it integrates advanced agentic capabilities with both instant and thinking response modes. The model supports conversational and agentic paradigms, enabling it to handle complex tasks that require sequential reasoning and tool use. Notably, its Agent Swarm technology allows coordination of up to 100 specialized AI agents working simultaneously—a parallel execution model that reduces processing time by 4.5x compared to sequential approaches.

The model emerged from Moonshot AI and quickly gained traction in production environments: Cursor acknowledged building its Composer 2 coding assistant directly on Kimi K2.5, and Alibaba Cloud incorporated it into its Coding Plan alongside other open-source models. On Humanity's Last Exam, Kimi K2.5 achieves 50.2% accuracy at roughly 76% lower cost than comparable proprietary models, making it a practical choice for teams seeking open-weight agentic performance. Available through platforms like Nscale for fully managed inference, the model balances frontier-level reasoning with cost efficiency, positioning it as a viable foundation for developers building autonomous multi-agent systems.

Alibaba Coding Plankimi-k2.5kimi-k2

Quick Info

Powered by
Provider
Alibaba Coding Plan
Model key
kimi-k2.5
Release date
Jan 27, 2026
Last updated
Jan 27, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Latest news about Kimi K2.5

Alibaba Coding Plan

CoverageBenchmark

CoreWeave Inference achieves the highest output speed for newly-launched Kimi K2.7 Code and ranks in the most attractive price-performance quadrant., Stream

Alibaba Coding Plan

CoverageBenchmark

Kimi K2.7-Code claims 30% fewer thinking tokens and a drop-in API swap path, but independent benchmarks show kernel regressions and no DeepSWE submission.

Alibaba Coding Plan

CoverageBenchmark

CoreWeave Inference achieves the highest output speed for Kimi K2.6 and ranks in the most attractive price-performance quadrant.

Alibaba Coding Plan

CoverageBenchmark

Cursor's Composer 2.5 undercuts Opus 4.7 and GPT-5.5 on price, posts gains on Terminal-Bench and SWE-Bench, but real-world coding tests loom.

Alibaba Coding Plan

CoverageRelease Notes

Open-source push continues with Kimi K2.6, but pressure to monetise is nudging some Chinese AI firms behind closed doors.

Alibaba Coding Plan

Coverage

Kimi K2.5 brings advanced agentic reasoning, design-to-code capabilities, and production-ready performance to fully managed inference endpoints on Nscale. Follow our blog for more insights from Nscale.

Alibaba Coding Plan

Coverage

The US AI startup Cursor has acknowledged that its newly introduced coding model Composer 2 is built on the Chinese open-source language model Kimi K2.5

Alibaba Coding Plan

Coverage

Alibaba Cloud has launched a new Coding Plan featuring four open-source model API services: Qwen3.5, GLM-5, MiniMax M2.5, and Kimi K2.5. According to PANews, this initiative makes Alibaba Cloud the on

Videos about Kimi K2.5

More models around Kimi K2.5