Alibaba Coding Plan
CoreWeave Inference achieves the highest output speed for newly-launched Kimi K2.7 Code and ranks in the most attractive price-performance quadrant., Stream
Model details
Kimi K2.5 is an open-source native multimodal agentic model designed to unify vision and language understanding within a single architecture. Built through continual pretraining on approximately 15 trillion mixed visual and text tokens atop the Kimi-K2-Base foundation, it integrates advanced agentic capabilities with both instant and thinking response modes. The model supports conversational and agentic paradigms, enabling it to handle complex tasks that require sequential reasoning and tool use. Notably, its Agent Swarm technology allows coordination of up to 100 specialized AI agents working simultaneously—a parallel execution model that reduces processing time by 4.5x compared to sequential approaches.
The model emerged from Moonshot AI and quickly gained traction in production environments: Cursor acknowledged building its Composer 2 coding assistant directly on Kimi K2.5, and Alibaba Cloud incorporated it into its Coding Plan alongside other open-source models. On Humanity's Last Exam, Kimi K2.5 achieves 50.2% accuracy at roughly 76% lower cost than comparable proprietary models, making it a practical choice for teams seeking open-weight agentic performance. Available through platforms like Nscale for fully managed inference, the model balances frontier-level reasoning with cost efficiency, positioning it as a viable foundation for developers building autonomous multi-agent systems.
A provider subscription or plan supersedes token-based pricing for this model.
Alibaba Coding Plan
CoreWeave Inference achieves the highest output speed for newly-launched Kimi K2.7 Code and ranks in the most attractive price-performance quadrant., Stream
Alibaba Coding Plan
Kimi K2.7-Code claims 30% fewer thinking tokens and a drop-in API swap path, but independent benchmarks show kernel regressions and no DeepSWE submission.
Alibaba Coding Plan
CoreWeave Inference achieves the highest output speed for Kimi K2.6 and ranks in the most attractive price-performance quadrant.
Alibaba Coding Plan
Cursor's Composer 2.5 undercuts Opus 4.7 and GPT-5.5 on price, posts gains on Terminal-Bench and SWE-Bench, but real-world coding tests loom.
Alibaba Coding Plan
Open-source push continues with Kimi K2.6, but pressure to monetise is nudging some Chinese AI firms behind closed doors.
Alibaba Coding Plan
Kimi K2.5 brings advanced agentic reasoning, design-to-code capabilities, and production-ready performance to fully managed inference endpoints on Nscale. Follow our blog for more insights from Nscale.
Alibaba Coding Plan
The US AI startup Cursor has acknowledged that its newly introduced coding model Composer 2 is built on the Chinese open-source language model Kimi K2.5
Alibaba Coding Plan
Alibaba Cloud has launched a new Coding Plan featuring four open-source model API services: Qwen3.5, GLM-5, MiniMax M2.5, and Kimi K2.5. According to PANews, this initiative makes Alibaba Cloud the on