Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Together AI logo

Model details

DeepSeek V4 Pro

DeepSeek V4 Pro is a 1.6-trillion parameter Mixture-of-Experts model that activates 49 billion parameters per forward pass, designed for highly efficient million-token context intelligence. The architecture incorporates a hybrid attention mechanism combining Compressed Sparse Attention and Heavily Compressed Attention, which reduces inference computational requirements to just 27% of the previous generation while cutting KV cache usage by 90%. Manifold-Constrained Hyper-Connections strengthen conventional residual connections for stable signal propagation across the massive model. Built to rival frontier closed models, this flagship targets advanced reasoning, complex software engineering tasks, and long-running agentic workflows where extended context understanding matters most.

The model launched in April 2026 under MIT open weights alongside a lighter Flash variant, continuing DeepSeek's pattern of releasing preview checkpoints for community evaluation. It reportedly beats all rival open models in mathematics and coding benchmarks, trailing state-of-the-art closed systems by only three to six months while costing a fraction of competitors' prices. With dual Thinking modes and a million-token default context window, V4-Pro suits developers building agents, research pipelines, or cost-sensitive production systems that need strong reasoning without proprietary lock-in. The combination of aggressive pricing, open access, and near-frontier performance positions it as a practical choice for teams seeking open-weight power at scale.

Together AIdeepseek-ai/DeepSeek-V4-Prodeepseek

Quick Info

Powered by
Provider
Together AI
Model key
deepseek-ai/DeepSeek-V4-Pro
Release date
Apr 24, 2026
Last updated
Apr 24, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.74
Output token cost
$3.48

Limits

Output tokens
384,000 tokens
Context window
512,000 tokens

Transparent token rates

Compare DeepSeek V4 Pro pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek V4 Pro

Together AI

Official sourceAnnouncement

Together AI announced on April 29, 2026 that DeepSeek V4 Pro is now available on its platform with a 512K-token context window for long-context reasoning workloads. The model uses a 1.6T-parameter MoE architecture with 49B activated parameters and exposes three controllable reasoning modes (Non-Think, Think High, Think Serverless pricing is published at $2.10 per 1M input tokens, $0.20 per 1M cached input tokens, and $4.40 per 1M output tokens, with a path to dedicated infrastructure for full 1M context and reserved capacity. The post notes DeepSeek V4 Flash is coming soon as a faster, lower-cost V4 option.

Nebius Token Factory

Coverage

DeepSeek officially released its V4-Pro model (general availability build designated DeepSeek-V4-Pro-0813) on August 13, 2026, after a preview period that began in April, making the model available across DeepSeek's app, web interface, and API. According to the supplied QZ coverage citing DeepSeek's own disclosures, th The same QZ report notes that a price increase for the V4 model family takes effect at 16:00 UTC on August 16, 2026, with V4-Pro output tokens rising to $3.96 per million at peak hours from the prior flat rate of $0.87 per million, and that DeepSeek is introducing peak and off-peak billing with off-peak rates at half t

Together AI

Official sourceBenchmark

DeepSeek V4 Pro 0813 is the official GA release of DeepSeek's flagship model on Together AI, superseding the earlier V4 Pro preview with substantially enhanced agentic capabilities. It retains the 1.6T total / 49B active MoE structure and now ships with a DSpark speculative decoding module attached for faster generatio Reasoning effort is adjustable per request across low, high, and max levels on the same deployment, letting a single endpoint serve fast responses and deep deliberation. DeepSeek's published benchmark suite shows the 0813 release improving on the preview across the board, with the largest gains on repository-scale engi

Together AI

CoverageBenchmark

DeepSeek V4 Pro Benchmark Review: From Parameter Race to Real‑World Task Fit After months of anticipation, DeepSeek officially released DeepSeek V4 on April 24, announcing that 1‑million‑token …

Together AI

Coverage

DeepSeek just launched its fourth generation of flagship models with DeepSeek-V4-Pro and DeepSeek-V4-Flash, both targeted at enabling highly efficient million…

Together AI

Coverage

Chinese startup says DeepSeek-V4-Pro beats all rival open models for maths and coding.

Together AI

Coverage

According to @deepseek_ai, the DeepSeek API now supports the new deepseek-v4-pro and deepseek-v4-flash models with 1M context windows and dual Thinking and...

Hugging Face

Coverage

DeepSeek's official Change Log entry dated August 13, 2026 confirms the GA rollout of DeepSeek-V4-Pro across the app, web, and API, with the calling method unchanged at model name deepseek-v4-pro. The GA release reports materially enhanced agent capabilities, including Terminal Bench 2.1 at 87.9, Cybergym at 83.3, Tool The same changelog documents native support for the OpenAI Responses API, specifically adapted for Codex with a one-click configuration script, and introduces three thinking-effort levels (low / high / max) for both V4-Pro and V4-Flash. The DeepSeek API model name deepseek-v4-pro now points at the GA build with substan

Together AI

Official sourceOfficial

1.6T parameter (49B activated) MoE model with 1M token context, hybrid attention requiring only 27% inference FLOPs and 10% KV cache vs V3.2, three reasoning modes, and 93.5% LiveCodeBench.

Videos about DeepSeek V4 Pro

More models around DeepSeek V4 Pro