Alibaba (China)
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It...
Model details
DeepSeek V3.1 is a large hybrid reasoning model from the DeepSeek family, with 671B total parameters and 37B active parameters. Its defining design choice is the ability to operate in either a thinking or non-thinking mode, selected through prompt templates, which lets developers choose between deeper deliberation and faster, more direct responses from the same underlying model. A container variant called DeepSeek-V3.1-Terminus also appears in the NVIDIA NGC catalog under the deepseek-ai organization, suggesting an updated or refreshed build of the base release that is packaged for enterprise deployment on NVIDIA infrastructure.
In practical terms, the model is aimed at workloads that benefit from adjustable reasoning depth, such as complex analytical tasks where chain-of-thought helps, alongside routine generation where low latency matters. The MoE-style activation pattern (37B active out of 671B) points to an architecture intended to balance capability with inference efficiency, and the availability of a managed NVIDIA container makes it approachable for teams that prefer not to self-host from scratch. Independent community commentary treats V3.1 as a recognized and noteworthy step in DeepSeek's roadmap, positioned between earlier generations and anticipated future releases, making it a reasonable choice for teams wanting a capable reasoning model with explicit control over how much deliberation it performs per request.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Alibaba (China)
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It...
This exact model name is also listed by 17 other providers.