Sulat.com
AI models
ai& logo

Model details

DeepSeek V4 Pro

DeepSeek V4 Pro is a large-scale Mixture-of-Experts language model built on 1.6 trillion total parameters with 49 billion activated per token, allowing it to combine broad knowledge capacity with comparatively light per-request compute. Its defining architectural choice is a hybrid attention design that pairs Compressed Sparse Attention with Heavily Compressed Attention, which the technical documentation reports reduces single-token inference FLOPs to roughly 27% of what DeepSeek V3.2 used at million-token context lengths. The same architecture underpins DeepSeek V4 Flash, positioning V4 Pro as the higher-capacity sibling tuned for more demanding workloads. This attention pairing, together with the 1M-token context window, is the core efficiency story for the model.

Post-training follows a deliberate two-stage pipeline in which separate domain experts are first cultivated through supervised fine-tuning and GRPO reinforcement learning, and then merged into a single model via on-policy distillation. The result is aimed squarely at advanced reasoning, coding, and long-horizon agent workflows such as full-codebase analysis, multi-step automation, and large-scale information synthesis, with configurable reasoning effort levels up to a maximum setting. Because the model is openly distributed on Hugging Face under an MIT license, teams can self-host for full-stack coding assistants, research pipelines, or enterprise agents that need both very long context and controllable reasoning depth, while still being able to call hosted endpoints for production traffic.

ai&deepseek-ai/deepseek-v4-prodeepseek-thinking

Quick Info

Powered by
Provider
ai&
Model key
deepseek-ai/deepseek-v4-pro
Release date
Apr 24, 2026
Last updated
Apr 24, 2026
Knowledge cutoff
2025-05
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$2.50

Limits

Output tokens
384,000 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare DeepSeek V4 Pro pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek V4 Pro

Hugging Face

Coverage

DeepSeek's official Change Log entry dated August 13, 2026 confirms the GA rollout of DeepSeek-V4-Pro across the app, web, and API, with the calling method unchanged at model name deepseek-v4-pro. The GA release reports materially enhanced agent capabilities, including Terminal Bench 2.1 at 87.9, Cybergym at 83.3, Tool The same changelog documents native support for the OpenAI Responses API, specifically adapted for Codex with a one-click configuration script, and introduces three thinking-effort levels (low / high / max) for both V4-Pro and V4-Flash. The DeepSeek API model name deepseek-v4-pro now points at the GA build with substan

Videos about DeepSeek V4 Pro

More models around DeepSeek V4 Pro