Sulat.com
AI models
OpenCode Go logo

Model details

Kimi K3

Kimi K3 is delivered through OpenCode Go's low-cost subscription tier, a service that curates popular open coding models and hosts them across US, EU, and Singapore regions for stable international access. The model is positioned as part of a curated set that includes peers like GLM-5.2 and Grok 4.5, selected by OpenCode after testing models directly with their teams and benchmarking provider combinations. Its open-weights status means developers can also self-host the weights themselves, while the OpenCode Go route offers managed, low-latency API access for those who prefer a turnkey setup.

The model handles text, image, and video inputs and produces text outputs, with a generous context window suited to long coding sessions and large repository analysis. On OpenCode's usage dashboard, Kimi K3 currently sits well behind flagship peers like DeepSeek V4 Flash and GLM-5.2 in raw token volume, attracting roughly 5B tokens and around 5,200 unique users in the observed window, with US, Chinese, and Japanese developers making up the largest share of activity. Practical sessions tend to average around 914K tokens at about $0.67 each, and a 91% cache-hit ratio on input tokens suggests the model is being used efficiently with substantial prompt caching, making it a reasonable fit for cost-sensitive workflows that benefit from long context and multimodal grounding.

OpenCode Gokimi-k3kimi-k3

Quick Info

Powered by
Provider
OpenCode Go
Model key
kimi-k3
Release date
Jul 16, 2026
Last updated
Jul 16, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$3.00
Output token cost
$15.00

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Latest news about Kimi K3

OpenCode Go

Official sourceDocumentation

OpenCode's official Go documentation lists Kimi K3 among the models available through the OpenCode Go subscription, alongside Grok 4.5, GLM-5.x variants, MiniMax M3, Qwen3.x, DeepSeek V4, and others. It is presented as a low-cost subscription — $5 for the first month, then $10/month — designed to give reliable global a Usage limits on OpenCode Go are defined in dollar terms rather than request counts: $12 per 5-hour window, $30 weekly, and $60 monthly, with the actual request count depending on the model used. The documentation notes that the model list may change as new models are tested and added, and instructs users to run /models

OpenCode Go

CoverageBenchmark

Yotta Labs' technical breakdown confirms Kimi K3 as Moonshot AI's 2.8-trillion-parameter open-weight release, announced mid-July 2026 with API access first and weights shipped July 27, 2026 as promised. Architecture is a sparse Mixture-of-Experts transformer with 896 experts and 16 active per token, using Kimi Delta At The page notes K3 does not fit on a single GPU or node, requiring a multi-node cluster with 1.56 TB+ of aggregate GPU memory, and frames the bigger story as routing between K3 and a frontier model outperforming either alone. Moonshot's API pricing is $3 per 1M input tokens ($0.30 cached) and $15 per 1M output, with str

OpenCode Go

CoverageBenchmark

This third-party Substack report (July 16, 2026) confirms Kimi K3 is live via Moonshot's API, kimi.com, mobile apps, Kimi Work, and Kimi Code, with OpenRouter providing immediate access for users without a Moonshot account. The model is described as a 2.8 trillion-parameter multimodal system with native vision, a 1,048 Independent benchmark coverage cited in the article places Kimi K3 at a score of 57 on the Artificial Analysis Intelligence Index, ranking it fourth among 189 models, behind Claude Fable 5 and two GPT-5.6 Sol reasoning settings and ahead of Claude Opus 4.8, GPT-5.5 at xhigh, Claude Sonnet 5, and GLM-5.2. An open-weight

OpenCode Go

Official sourceOfficial

OpenCode's usage data page for Kimi K3 ranks the model 15th in recent OpenCode Go usage, capturing roughly 5B tokens and about 0% of an observed 2M-volume window, with 5.2K unique users and 5,513 completed sessions. Listed specifications include a 1M-token context window, up to 131K output tokens, text/image/video inpu The page also reports an average session cost of about $0.67, average tokens per session of 914K, and a 91% cache ratio on input tokens. Geographic breakdown shows the United States at roughly 1B tokens (20%) leading usage, followed by China (700M, 15%), Japan (400M, 8%), Brazil, Singapore, India, Mexico, and others. P

OpenCode Go

CoverageBenchmark

Simon Willison's July 16, 2026 weblog post reports that Moonshot AI announced Kimi K3 as its "most capable model to date, with 2.8 trillion parameters," branding it the first "open 3T-class model" and surpassing DeepSeek's 1.6T v4 Pro, with an open-weight release promised by July 27, 2026. Artificial Analysis highlight Willison flags Kimi K3's $3/M input and $15/M output pricing as the most expensive release from a Chinese AI lab to date, matching Anthropic's Claude Sonnet tier and a sharp increase over Kimi K2.6's $0.95/$4, alongside the 2.8T-parameter jump from K2.6's ~1T. He also documents a practical OpenRouter test where K3 spen

OpenCode Go

Official sourceBenchmark

OpenCode's aggregate data dashboard, last updated July 16, 2026, tracks real token usage across OpenCode Go models. Total tracked volume has grown from roughly 1.5T tokens in late May 2026 to peaks above 3.3T in early July, with daily figures on July 16, 2026 landing around 2.8T tokens. The dashboard ranks models by re Top models by usage include DeepSeek V4 Flash at 11T (+5%), DeepSeek V4 Pro at 3.9T (+1%), GLM-5.2 at 1.3T (−17%), MiMo-V2.5 at 1.0T (−8%), and MiniMax M3 at 847B (−23%). Kimi K3 sits further down the rankings at 5B tokens alongside Kimi K2.5 (6B), with several Moonshot, Zhipu, Qwen, MiniMax, and Tencent models populat

Videos about Kimi K3

More models around Kimi K3