Currently listed through these providers:
Model details
Kimi K3
Kimi K3 is introduced as the provider's most capable model to date, a 2.8-trillion-parameter system that the official announcement describes as the world's first open 3T-class model. It is built on two named architectural innovations, Kimi Delta Attention and Attention Residuals, paired with native vision capabilities and a 1-million-token context window. The provider's own framing positions it for frontier intelligence across long-horizon coding, knowledge work, and reasoning, while acknowledging that its overall benchmark performance still trails the leading proprietary models, identified in the announcement as Claude Fable 5 and GPT 5.6 Sol, even as it consistently outperforms other tested models within the provider's evaluation suite.
In practical deployment, Kimi K3 is exposed through the Kimi Code product under two model-ID variants. The full-context "k3 1M" option targets the longest tasks and consumes roughly twice the quota of the trimmed "k3-256k" variant, which is recommended for everyday Q&A, code completion, routine feature development, and single-file or small-file edits where video input is not required. This makes the smaller variant a cost-efficient default for common coding workflows while reserving the longer-context configuration for workloads needing a million-token reasoning horizon, aligning the model's architecture and context design with the realities of agentic development sessions.
Quick Info
Powered by- Provider
- Neon
- Model key
- kimi-k3
- Release date
- Jul 16, 2026
- Last updated
- Jul 16, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $3.00
- Output token cost
- $15.00
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Kimi K3
Videos about Kimi K3
More models around Kimi K3
This exact model name is also listed by 60 other providers.