Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Azure logo

Model details

DeepSeek-V4-Pro

DeepSeek-V4-Pro is built in the deepseek-thinking family as a Mixture-of-Experts model with roughly 1.6 trillion total parameters and about 49 billion active per token, a design that keeps inference work lean while preserving a very large knowledge base. The model ships with widely accessible open weights, encouraging local deployment, fine-tuning, and transparency research rather than locking capabilities behind a closed API. Its training direction emphasizes deep reasoning, agentic coding, and rich world knowledge, aiming to behave like a careful, step-by-step thinker rather than a fast conversational assistant.

At Preview release, DeepSeek-V4-Pro was highlighted as open-source state-of-the-art on agentic coding benchmarks and as leading current open models on math, STEM, and general coding tasks, trailing only a leading closed-source frontier model on broad world knowledge. The release paired those results with a very long context window marketed as a cost-effective the cataloged API limit token window, making it well suited to repository-scale code analysis, multi-document research, and long agent traces. Practically, it fits teams that need strong reasoning and open deployment without per-token ceiling pressure, while a smaller sibling in the same family offers a faster, cheaper alternative for simpler tasks.

Azuredeepseek-v4-prodeepseek-thinking

Quick Info

Powered by
Provider
Azure
Model key
deepseek-v4-pro
Release date
Apr 24, 2026
Last updated
Apr 24, 2026
Knowledge cutoff
2025-05
AI SDK package
@ai-sdk/openai-compatible
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.74
Output token cost
$3.48

Limits

Output tokens
384,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare DeepSeek-V4-Pro pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek-V4-Pro

Ollama Cloud

CoverageBenchmark

Aikido's security-focused benchmark study (published August 21, 2026) burned roughly 11.7 billion tokens testing ten contenders across three attempts on 32 fresh, off-the-shelf vulnerabilities. The lineup added GLM-5.3, DeepSeek V4 Pro 0813, DeepSeek V4 Flash 0731, Qwen3.8-Max, Kimi K3, and Grok 4.6 to its evaluation, The study concludes that open-source models now outperform the public closed frontier on pooled vulnerability recall, with DeepSeek V4 Pro topping every public closed model tested and Qwen, Kimi, and GLM-5.3 offering strong consistency. It also flags a trade-off: while coverage rose, the cheaper open models produced th

Azure

Coverage

Quartz's article reports the official general-availability launch of DeepSeek-V4-Pro-0813 on Thursday, August 13, 2026, following an April preview period. It states that the GA checkpoint focuses on agent capabilities — tool use, code execution, and multi-step workflows — and lists DeepSeek's own benchmark numbers: Ter The piece additionally documents a pricing restructure effective 16:00 UTC on August 16, 2026: V4-Pro output tokens were set to rise to $3.96 per million at peak hours from a prior flat rate of $0.87 per million, alongside a new peak/off-peak billing split. These are model-level capability and pricing facts drawn from

Azure

CoverageBenchmark

The release-tracker page documents DeepSeek-V4-Pro-0813, released on August 13, 2026, as a model-level update from DeepSeek itself. It enumerates the benchmark suites used to evaluate the model — BullshitBench v2, SWE-Bench Verified, Terminal-Bench 2.1, Toolathlon-Verified, BrowseComp, CyberGym, and two additional test The page is centered on the model itself rather than on any serving provider, which is appropriate for a model-focused subject. It also lists DeepSeek as the operator running the model and answering requests, with context-window and provider-availability fields that go beyond pure pricing. There is no Azure-specific ho

Azure

CoverageRelease Notes

Vercel's AI Gateway changelog (June 11, 2026) states explicitly that "Azure is now a provider for DeepSeek V4 Pro and V4 Flash on AI Gateway," with SDK identifiers deepseek/deepseek-v4-pro and deepseek/deepseek-v4-flash. Requests route through Azure automatically with no code changes, and customers can use BYOK with ex This is the cleanest indirect confirmation that Azure exposes DeepSeek-V4-Pro as a callable model, though it is a Vercel-layer announcement rather than an Azure-native model card. It does not specify Azure region, SKU, or quota specifics, but it does validate that DeepSeek-V4-Pro is reachable through Azure in productio

Azure

CoverageRelease Notes

Microsoft's "What's new in Microsoft Foundry | May 2026" devblog confirms that the DeepSeek V4 family expands open-model choice in the Foundry catalog, and separately notes that "Fireworks AI — May update: DeepSeek V4 Pro and Kimi 2.6 arrive via Fireworks for high-performance open-model inference." This is the closest The same Foundry post links DeepSeek-V4-Pro to Fireworks as a delivery path on Azure rather than a first-party Azure SKU, so it does not confirm a native Azure model deployment entry. Readers should treat the Azure availability as catalog/Fireworks-routed rather than a fully Azure-managed deployment. Other May 2026 Fou

Azure

Coverage

The DeepSeek API change log dated September 10, 2026 announces DeepSeek-V4.1-Flash as the smallest model in a new architecture family with native multimodal visual understanding, and explicitly states that V4.1 Flash now outperforms DeepSeek V4 Pro across performance, cost, speed, and total time. It documents a concret Critically for the DeepSeek-V4-Pro subject, the changelog announces an orderly retirement of V4 Pro: after 12:00 Beijing Time on September 14, 2026, and until the future release of V4.1 Pro, all requests to deepseek-v4-pro will be routed to V4.1 Flash and billed at V4.1 Flash pricing. This is a model-level behavioral c

Videos about DeepSeek-V4-Pro

More models around DeepSeek-V4-Pro