Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Melious logo

Model details

DeepSeek V4 Pro

DeepSeek V4 Pro is the flagship of the V4 family introduced by DeepSeek as an open-sourced preview release, built around a mixture-of-experts design with 1.6T total parameters and 49B active parameters per inference. The model was positioned at launch as rivaling the strongest closed-source systems, while its open weights and accompanying technical report are published on Hugging Face under the DeepSeek organization, making it accessible for self-hosting and research. A smaller sibling, DeepSeek V4 Flash, was released the same day as a faster, lower-cost alternative, indicating that V4 Pro anchors a deliberately tiered family aimed at different latency and budget trade-offs.

The model's intended strengths center on extended reasoning and knowledge-intensive tasks: DeepSeek highlights enhanced agentic coding capabilities where it claimed open-source state-of-the-art results at release, alongside broad world knowledge where it was reported to lead other open models and approach the top closed competitors. Its very long context window enables agents and research workflows that require sustained reasoning across large documents, and the MoE architecture keeps active compute modest relative to total capacity, supporting cost-efficient inference at scale. A later DeepSeek announcement explicitly treats V4 Pro as the flagship benchmark target, reinforcing its role as the high-end reference point in the lineup for complex reasoning, coding, and agentic applications.

Meliousdeepseek-v4-prodeepseek-thinking

Quick Info

Powered by
Provider
Melious
Model key
deepseek-v4-pro
Release date
Apr 24, 2026
Last updated
Apr 24, 2026
Knowledge cutoff
2025-05
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.85472
Output token cost
$3.70944

Limits

Output tokens
384,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare DeepSeek V4 Pro pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek V4 Pro

Melious

Coverage

DeepSeek officially released the V4-Pro model on Thursday, August 13, 2026, making the general-availability version designated DeepSeek-V4-Pro-0813 available across its app, web interface, and API after a preview period that began in April 2026. The GA release focuses on agent capabilities โ€” tasks where AI systems use The V4-Pro API was updated to work with the OpenAI Responses API format out of the box and includes built-in support for Codex integration, while thinking effort levels for both V4-Pro and V4-Flash were expanded to three settings โ€” low, high, and max โ€” allowing developers to match computation intensity to task complexi

Melious

CoverageBenchmark

BenchLM's third-party model profile for DeepSeek V4 Pro 0813 (released Aug 13, 2026) records a composite capability score of 66.4/100, ranking the model 32nd out of 232 tracked models as of September 10, 2026. The profile reports 40 source-displayable benchmark rows, with Reasoning ranking 15th as the strongest eligibl The aggregator lists API pricing at $0.435 input and $0.87 output per million tokens, with cached input at $0.003625 per million and a blended figure of $0.65, alongside a reported 1M-token context window, 69 tok/s speed, and 30.68 s first-token latency (field medians of 86.5 tok/s and 256,000 tokens respectively). The

Melious

Coverage

DeepSeek's official API changelog dated 2026-09-10 announces the planned orderly retirement of DeepSeek V4 Pro. The changelog states that extensive internal testing shows V4.1 Flash now outperforms DeepSeek V4 Pro across performance, cost, speed, and total time, so DeepSeek plans to retire V4 Pro in an orderly manner. The 2026-09-10 entry is part of the same official changelog that introduces DeepSeek-V4.1-Flash, a new architecture family member with native multimodal visual understanding designed for higher capability ceiling, faster inference, and higher throughput. V4.1 Flash can be called via the model name "deepseek-flash" and

Videos about DeepSeek V4 Pro

More models around DeepSeek V4 Pro