Azure
DeepSeek-V3.2-Speciale is a high-compute variant of DeepSeek-V3.2 optimized for maximum reasoning and agentic performance. 131,072 token context window. Includes independent benchmarks from Artificial Analysis.
Model details
DeepSeek-V3.2-Speciale sits at the top of the DeepSeek-V3.2 family as a high-compute variant engineered specifically for the most demanding reasoning and agentic workloads. The architecture builds on DeepSeek Sparse Attention (DSA) for efficient long-context processing, enabling the model to maintain coherent performance across extended reasoning chains and multi-turn interactions. Where the standard V3.2 balances inference speed with capability, Speciale leans fully into maximum reasoning performance—pushing post-training reinforcement learning to greater intensity to unlock higher capability ceilings. This design makes it particularly suited for tasks requiring deep deliberation, complex mathematical problem-solving, and competitive programming scenarios where benchmark dominance matters.
The model benefits from a large-scale agentic task synthesis pipeline that generates training data across more than 1,800 simulated environments and 85,000 complex instruction scenarios, substantially improving compliance and generalization in interactive settings. Speciale represents the first DeepSeek model to integrate thinking directly into tool-use, though it currently excludes tool-calling from its API interface to support open community evaluation and research. Evaluations report gold-medal-level results in IMO, CMO, ICPC World Finals, and IOI 2025 competitions, placing it ahead of GPT-5 on difficult reasoning benchmarks and on par with Gemini-3.0-Pro—while retaining strong coding and tool-use reliability for production agent workflows.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Azure
DeepSeek-V3.2-Speciale is a high-compute variant of DeepSeek-V3.2 optimized for maximum reasoning and agentic performance. 131,072 token context window. Includes independent benchmarks from Artificial Analysis.
Azure
@unsloth @ shimmyshimmer is it DSA what's slowing down release? I'd love to help!