Currently listed through these providers:
Model details
DeepSeek V3 0324
DeepSeek V3 0324 is a large open-weights mixture-of-experts language model built as a refined evolution of the original DeepSeek V3. With around 685 billion total parameters, it is designed to push deeper analytical reasoning while keeping generation efficient enough for practical deployment, and its structure stays compatible with the earlier V3 release so existing tooling and local setups can be reused. The model is positioned as a general-purpose workhorse that handles long-context text across both English and Chinese, with explicit tuning toward higher-quality prose, code, and step-by-step problem solving rather than narrow single-task behavior.
Compared with its predecessor, the 0324 refresh delivers clear benchmark gains in reasoning-heavy evaluations, including jumps on MMLU-Pro, GPQA, AIME, and LiveCodeBench, alongside cleaner front-end web code and more natural Chinese long-form writing. It is recommended for applications that need strong analytical chains, reliable function calling, and polished Chinese output, and it supports structured features like JSON output and Fill-in-the-Middle completion that fit production development pipelines. As an open-weights model optimized for inference throughput, it slots naturally into agent, retrieval, and coding workflows that need reasoning depth without moving to a closed proprietary system.
Quick Info
Powered by- Provider
- Meganova
- Model key
- deepseek-ai/DeepSeek-V3-0324
- Release date
- Mar 24, 2025
- Last updated
- Mar 24, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.25
- Output token cost
- $0.88
Limits
- Output tokens
- 163,840 tokens
- Context window
- 163,840 tokens
Latest news about DeepSeek V3 0324
No articles yet. Fetch the latest news to show it here.