Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

Kimi K2 0905

Kimi K2 0905 is an open-weight update to Moonshot AI's K2 family, specifically refined for agentic coding and developer workflows. It brings three headline advances over the July checkpoint: a doubled context window now reaching 262k tokens, substantially improved tool-calling reliability, and stronger front-end code generation. The model's attention mechanism was explicitly tuned for long-context scenarios, maintaining coherence across the full window so developers can work with larger codebases, conversation histories, and test suites without the typical degradation at context boundaries. Its prior generation already earned attention for consistent diff generation at roughly 5%, on par with leading closed models, and the 0905 release builds on that foundation with sharper focus on the capabilities that matter most for agent-driven development.

Beyond raw coding ability, the practical deployment story matters: Kimi K2 0905 is served across multiple providers including GroqCloud, where prompt caching unlocks up to 50% savings on cached tokens, making long-context agent loops economically viable. Groq's infrastructure delivers this model at around 349 tokens per second, handling production workloads without throttling and removing model latency as a workflow bottleneck. The combination of a very large context window, reliable tool calling, and competitive front-end generation makes Kimi K2 0905 a strong fit for teams building coding assistants, multi-turn agent pipelines, and developer tools that need sustained context retention over extended sessions.

NovitaAImoonshotai/kimi-k2-0905kimi-k2

Quick Info

Powered by
Provider
NovitaAI
Model key
moonshotai/kimi-k2-0905
Release date
Sep 5, 2025
Last updated
Sep 5, 2025
Knowledge cutoff
2024-10
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.60
Output token cost
$2.50

Limits

Output tokens
98,304 tokens
Context window
262,144 tokens

Transparent token rates

Compare Kimi K2 0905 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Kimi K2 0905

NovitaAI

CoverageRelease Notes

Opper AI's Moonshot release tracker confirms Kimi K2 0905 was released on 5 September 2025 with a 262K token context window, listed input pricing of $0.60 per million tokens and output pricing of $2.50 per million tokens. The tracker places Kimi K2 0905 in the Moonshot release sequence between the original Kimi K2 and The tracker attributes an intelligence score of 15 to Kimi K2 0905 within its composite methodology, without detailing the scoring approach. Out of 7 tracked Moonshot releases shown on the page, 5 are noted as running on Opper, giving developers a consolidated view of the Kimi family's pricing and capability progressio

Jiekou.AI

CoverageBenchmark

MoonshotAI's Kimi K2 0905 is documented as a Mixture-of-Experts model with 1 trillion total parameters and 32 billion active parameters (rumoured), featuring a 262K token context window and a release date of September 4, 2025. Benchable reports a 100% success rate across all tracked benchmarks, perfect accuracy in Hall Pricing on the page is listed at $0.6 per 1M prompt tokens and $2.5 per 1M completion tokens, with Moonshot AI's own endpoint shown as the primary provider entry alongside additional routing endpoints. The model supports tools, response format, and structured outputs, aligning with Moonshot's positioning of the 0905 up

Videos about Kimi K2 0905

More models around Kimi K2 0905