DeepSeek V4 Pro 0813 marks the general availability release of DeepSeek's flagship Mixture-of-Experts system, graduating from preview status in mid-August 2026 with a refined build focused on agent-style engineering workloads. The architecture is reported at roughly 1.6 trillion total parameters with around 49 billion active per token, paired with a hybrid attention design that aims to keep inference costs manageable across very long contexts, and the model was pre-trained on more than 32 trillion tokens. It carries the same long-context backbone as the rest of the V4 family, including the smaller V4 Flash sibling, but pushes further into multi-step reasoning, tool use, and full-stack development tasks.
Beyond raw scale, the 0813 build is tuned for sustained agentic work, exposing reasoning effort controls, tool calling, and JSON-style structured outputs that let teams wire it into planning, verification, and code-execution loops. Independent reporting positions it near the top of contemporary agent benchmarks, including a reported 80.6% on SWE-bench Verified, with notable strengths on Terminal Bench 2.1, Cybergym, DeepSWE, and AutomationBench compared with leading frontier systems. The open-weight posture makes it attractive for organizations that want a high-capability reasoning and coding engine they can self-host, while third-party serving platforms continue to expose it across many regions for teams that prefer managed inference.