Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

Deepseek V3.1 Terminus

DeepSeek-V3.1-Terminus is a large hybrid reasoning model built on the DeepSeek-V3 foundation, designed to switch between deliberate analytical thinking and direct response modes depending on the task. The architecture extends the base model through a two-phase long-context training process, achieving up to 128K token context windows, while FP8 microscaling keeps inference efficient at scale. A reasoning_enabled boolean parameter gives users direct control over the model's thinking behavior, and the update cycle addressed user-reported issues around language consistency—significantly reducing the mixing of Chinese and English text and eliminating abnormal character glitches that appeared in earlier releases.

The model demonstrates measurable gains across reasoning and agentic benchmarks, with benchmark tables showing MMLU-Pro climbing from 84.8 to 85.0, GPQA-Diamond rising from 80.1 to 80.7, and notably Humanity's Last Exam jumping from 15.9 to 21.7, alongside BrowseComp reaching 38.5. Specialized agent modes for code execution and information retrieval position it well for development and research applications requiring reliable tool integration. Teams can further customize behavior through LoRA fine-tuning on supported deployment platforms, while the MIT license enables full open-weight access for redistribution and modification. On challenging reasoning tasks, it matches DeepSeek-R1 performance while delivering faster responses—a practical balance for production environments.

NovitaAIdeepseek/deepseek-v3.1-terminusdeepseek

Quick Info

Powered by
Provider
NovitaAI
Model key
deepseek/deepseek-v3.1-terminus
Release date
Sep 22, 2025
Last updated
Sep 22, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.27
Output token cost
$1.00

Limits

Output tokens
32,768 tokens
Context window
131,072 tokens

Transparent token rates

Compare Deepseek V3.1 Terminus pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Deepseek V3.1 Terminus

No articles yet. Fetch the latest news to show it here.

Videos about Deepseek V3.1 Terminus

More models around Deepseek V3.1 Terminus