LLM Gateway
DeepSeek says both models are more efficient and performant than DeepSeek V3.2 due to architectural improvements, and have almost "closed the gap" with current leading models, both open and closed, on reasoning benchmarks.
Model details
DeepSeek V3.2 is an open-source reasoning and agentic model family built to harmonize computational efficiency with strong performance across reasoning and tool-use tasks. At its core lies DeepSeek Sparse Attention (DSA), an architectural innovation that achieves fine-grained sparse attention to reduce computational complexity while maintaining quality in long-context scenarios. The model family includes a high-compute variant, DeepSeek-V3.2-Speciale, which pushes reasoning boundaries further to rival closed frontier models. Designed from the ground up for agents, this is the first DeepSeek model to integrate thinking directly into tool-use, supporting both thinking and non-thinking modes for flexible deployment across interactive environments.
The model leverages a scalable reinforcement learning framework with robust RL protocols and significant post-training compute scaling to achieve its reasoning capabilities. A defining strength is the large-scale agentic task synthesis pipeline, which systematically generated training data across more than 1,800 environments and 85,000-plus complex instructions, enabling substantial improvements in generalization and instruction-following robustness within interactive settings. The V3.2-Speciale variant has demonstrated gold-medal-level performance at the International Mathematical Olympiad, International Olympiad in Informatics, Chinese Mathematical Olympiad, and ICPC World Finals. Released under an MIT license, the model provides an open foundation for researchers and developers building advanced agentic applications, with its architecture optimized for long-context efficiency and its training methodology emphasizing reasoning depth over brute-force scale.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
LLM Gateway
DeepSeek says both models are more efficient and performant than DeepSeek V3.2 due to architectural improvements, and have almost "closed the gap" with current leading models, both open and closed, on reasoning benchmarks.
LLM Gateway
DeepSeek released DeepSeek-V3.2, a family of open-source reasoning and agentic AI models. The high compute version, DeepSeek-V3.2-Speciale, performs better than GPT-5 and comparably to Gemini-3.0-Pro
LLM Gateway
Updated March 2026: Comprehensive guide to DeepSeek V3.2, V4 (expected April 2026), R1/R2 reasoning models, and how to use DeepSeek in Antigravity via the OpenAI compatibility layer.
This exact model name is also listed by 27 other providers.