Currently listed through these providers:
Model details
deepseek-ai/DeepSeek-V4-Pro
DeepSeek-V4-Pro is a large open-weight language model from DeepSeek, designed for advanced reasoning, complex software engineering, and long-running agentic tasks. According to a third-party hosting page, an updated variant called DeepSeek-V4-Pro-0813 has been published as the official release, superseding an earlier preview, and is built on the same base model structure as that preview with a DSpark speculative decoding module attached to improve inference behavior. The framing emphasizes production-grade agentic capabilities, suggesting the model is targeted at workflows that chain tool use, planning, and multi-step problem solving rather than single-turn generation.
On the Fireworks AI hosting page, the DeepSeek-V4-Pro-0813 variant exposes a Playground for quick experimentation, an on-demand deployment option for dedicated capacity, and a fine-tuning flow that uses LoRA-based customization, giving teams multiple integration paths from prototyping to production. The speculative decoding component is positioned as a way to accelerate response generation without changing the underlying model outputs, which is useful for latency-sensitive interactive agents. Together, these traits indicate a practical fit for developers building tool-using assistants and code-focused applications that benefit from open weights and flexible deployment.
Quick Info
Powered by- Provider
- SiliconFlow (China)
- Model key
- deepseek-ai/DeepSeek-V4-Pro
- Release date
- Apr 24, 2026
- Last updated
- Apr 24, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.74
- Output token cost
- $3.48
Limits
- Output tokens
- 393,000 tokens
- Context window
- 1,049,000 tokens