Currently listed through these providers:
Model details
DeepSeek V4 Pro
DeepSeek V4 Pro is a very large Mixture-of-Experts model aimed squarely at agentic, code-heavy workloads, with 1.6T total parameters and 49B active on each forward pass. Host listings frame it as a continuation of the V4 line rather than a ground-up redesign: the 0813 served variant builds on the earlier Pro preview architecture and adds a DSpark speculative decoding module, which is the kind of pairing that typically helps latency and throughput without changing the model's core reasoning behavior. Its design priorities are visible in how it is positioned for advanced coding, tool use, terminal-style automation, and long-running agent tasks, rather than open-ended chat.
In practical terms, the model is shaped for sustained, multi-step work over large codebases and long documents, with one provider offering a 262K-token context window specifically for that use case. Users can dial reasoning depth through low, high, and max effort settings, which makes it adaptable to both quick lookups and deep, deliberation-heavy passes. Published benchmarks underline that focus: 87.9 on Terminal Bench 2.1, 42.7 / 60.0 on HLE without and with tools, and 61.5 on NL2Repo, putting it ahead of the V4 Flash and Pro preview variants on agentic and repository-level tasks. Together, the architecture, configurable reasoning, and benchmark profile point to a model best suited for teams building coding agents, automated pipelines, and other long-horizon AI systems rather than lightweight conversational assistants.
Quick Info
Powered by- Provider
- Model Oracle AI
- Model key
- deepseek-v4-pro
- Release date
- Apr 24, 2026
- Last updated
- Apr 24, 2026
- Knowledge cutoff
- 2025-05
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 384,000 tokens
- Context window
- 1,000,000 tokens
Latest news about DeepSeek V4 Pro
Videos about DeepSeek V4 Pro
More models around DeepSeek V4 Pro
This exact model name is also listed by 51 other providers.