Currently listed through these providers:
Model details
DeepSeek V4 Pro
DeepSeek V4 Pro is positioned as a flagship reasoning-tier release in DeepSeek's lineup, scaled as a 1.6-trillion parameter Mixture-of-Experts design that activates only a subset of its weights per token. That sparsity lets it carry an unusually wide knowledge base while keeping inference cost closer to a much smaller dense model, which is the core architectural story behind its push into long-running, multi-step work. It is distributed under an open MIT license, so the weights can be self-hosted, audited, or fine-tuned by teams that prefer to run their own stack rather than depend on a hosted API.
The model is aimed at workloads that benefit from extended deliberation and tool use, including complex software engineering, multi-file refactoring, agent-style task solving, and analyses that need to keep very large codebases or documents in working memory. Within the broader V4 family, DeepSeek has continued to invest in reasoning quality and tool-calling reliability, and a sibling multimodal release shows the family is also expanding into vision-grounded agent tasks. For practitioners, V4 Pro fits best when the work is heavy on chain-of-thought reasoning, structured output, and long-context retrieval rather than cheap, high-volume chat.
Quick Info
Powered by- Provider
- TokenGo
- Model key
- deepseek/deepseek-v4-pro
- Release date
- Apr 24, 2026
- Last updated
- Apr 24, 2026
- Knowledge cutoff
- 2025-05
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.435
- Output token cost
- $0.87
Limits
- Output tokens
- 384,000 tokens
- Context window
- 1,000,000 tokens
Latest news about DeepSeek V4 Pro
Videos about DeepSeek V4 Pro
More models around DeepSeek V4 Pro
This exact model name is also listed by 51 other providers.