Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Hugging Face logo

Model details

DeepSeek V4 Pro 0813

DeepSeek V4 Pro 0813 marks the general availability release of DeepSeek's flagship Mixture-of-Experts system, graduating from preview status in mid-August 2026 with a refined build focused on agent-style engineering workloads. The architecture is reported at roughly 1.6 trillion total parameters with around 49 billion active per token, paired with a hybrid attention design that aims to keep inference costs manageable across very long contexts, and the model was pre-trained on more than 32 trillion tokens. It carries the same long-context backbone as the rest of the V4 family, including the smaller V4 Flash sibling, but pushes further into multi-step reasoning, tool use, and full-stack development tasks.

Beyond raw scale, the 0813 build is tuned for sustained agentic work, exposing reasoning effort controls, tool calling, and JSON-style structured outputs that let teams wire it into planning, verification, and code-execution loops. Independent reporting positions it near the top of contemporary agent benchmarks, including a reported 80.6% on SWE-bench Verified, with notable strengths on Terminal Bench 2.1, Cybergym, DeepSWE, and AutomationBench compared with leading frontier systems. The open-weight posture makes it attractive for organizations that want a high-capability reasoning and coding engine they can self-host, while third-party serving platforms continue to expose it across many regions for teams that prefer managed inference.

Hugging Facedeepseek-ai/DeepSeek-V4-Pro-0813deepseek-thinking

Quick Info

Powered by
Provider
Hugging Face
Model key
deepseek-ai/DeepSeek-V4-Pro-0813
Release date
Aug 12, 2026
Last updated
Aug 22, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.32
Output token cost
$3.96

Limits

Output tokens
384,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare DeepSeek V4 Pro 0813 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek V4 Pro 0813

Together AI

CoverageBenchmark

An independent technical analysis published August 26, 2026 walks through the architecture behind DeepSeek-V4-Pro-0813, which shipped as a quiet general-availability release on August 13, 2026. The model introduces a hybrid attention architecture with new techniques for stabilizing very deep residual stacks, paired wit The blog places the 0813 release in context against competing reporting, citing an SCMP piece that found the updated V4-Pro struggles on certain general benchmarks while excelling at cybersecurity tasks. It frames the GA as a substantial engineering update — hybrid attention plus deep residual stabilization — rather th

Hugging Face

CoverageBenchmark

DeepSeek V4 Pro 0813 is an open-weight mixture-of-experts model released by DeepSeek under the MIT license, with the V4-Pro-0813 checkpoint going GA on August 13, 2026. Per the supplied page excerpt, it carries 1.6 trillion total parameters with 49 billion active per token, a 1-million-token context window with a 384K The same source documents self-hosting requirements and serving performance: an 8x B200 node (SGLang, NVFP4) or a 4x GB300 node (vLLM) is the reference configuration for V4 Pro, with InferenceX reporting a single-user throughput of 216 tok/s on a B300 node using SGLang FP4. Lambda pricing for an 8x B200 node is listed

Together AI

Coverage

DeepSeek officially launched its flagship model DeepSeek-V4-Pro-0813 on August 13, 2026, making it generally available across the company's app, web interface, and API after an April preview. The release focuses on agent capabilities — tasks where AI systems use tools, execute code, and complete multi-step workflows wi The launch introduced several developer-facing API updates: the V4-Pro API now works with the OpenAI Responses API format out of the box and includes built-in support for Codex integration. Thinking effort levels for both V4-Pro and V4-Flash have been expanded to three settings — low, high, and max — letting developers

Together AI

CoverageBenchmark

The release-tracker entry for DeepSeek-V4-Pro-0813 catalogs the model's August 13, 2026 general-availability launch, positioned as DeepSeek's flagship release 13 days after DeepSeek-V4-Flash-0731. The benchmark suite tracked includes BullshitBench v2, SWE-Bench Verified, Terminal-Bench 2.1, Toolathlon-Verified, BrowseC Pricing published by DeepSeek is reproduced at $0.66 per million input tokens, $0.022 per million cached input, and $1.98 per million output during off-peak hours, rising to $1.32 / $0.044 / $3.96 during peak windows of 01:00–04:00 and 06:00–10:00 UTC. The tracker notes that thinking and non-thinking modes are priced i

Hugging Face

CoveragePreview

GMICloud reports that DeepSeek V4 Pro’s 0813 build left preview on August 12, 2026, naming the exact DeepSeek-V4-Pro-0813 variant and describing it as a flagship model with a 1-million-token context window and 384,000-token maximum output. The post also identifies non-thinking, high-effort, and max-effort reasoning mod For developer impact, the post says the 0813 build targets agentic work and coding, with reported improvements in multi-step coding, tool use, cybersecurity workflows, and full-stack application development, based on evaluations using DeepSeek’s Harness framework. It also states that the model has 1.6 trillion total pa

Hugging Face

CoverageBenchmark

OpenRouter lists DeepSeek V4 Pro 0813 as a generally available release released on August 12, 2026, with a 1-million-token context window. The page identifies the subject as a large-scale mixture-of-experts model and displays routing, pricing, latency, throughput, and uptime information for multiple hosts. The exact variant name, GA status, release date, and context length provide corroborating metadata for the 0813 build. Because the page is primarily a gateway and hosting marketplace, its per-provider commercial and routing metrics are not treated as intrinsic model news; the product impact is limited to confirming tha

Hugging Face

CoverageRelease Notes

Fireworks AI’s changelog states that DeepSeek V4 Pro was deprecated from its serverless service effective August 27, 2026, and directs users to migrate to DeepSeek V4 Pro (0813). This is current, exact-version evidence that the 0813 build is the migration target on Fireworks’ platform. The notice is operationally useful for developers currently using the older Fireworks serverless model, because it identifies the required replacement version and the effective deprecation date. The changelog excerpt does not provide new model capabilities, benchmarks, or creator-issued launch details, so its impact is

Videos about DeepSeek V4 Pro 0813

More models around DeepSeek V4 Pro 0813