Sulat.com
AI models
Helicone logo

Model details

DeepSeek V3.1 Terminus

DeepSeek V3.1-Terminus is an iterative refinement of DeepSeek's V3 language model family, built to sharpen performance on tasks where the original V3.1 showed user-reported weaknesses. The model ships with hybrid thinking capabilities, letting developers toggle between reasoning and non-reasoning modes depending on the task at hand, and it is designed to function as a strong foundation for AI agents—particularly those built around code generation and web search workflows. Its architecture inherits the Mixture-of-Experts backbone and multi-head latent attention mechanisms that define the broader DeepSeek V3 series, while the updated Terminus checkpoint addresses language consistency issues such as unwanted Chinese-English mixing and stray abnormal characters that users flagged in earlier releases.

The Terminus update was driven by community feedback and refined through targeted optimizations across agentic benchmarks. Benchmark results show meaningful gains: Humanity's Last Exam scores climbed from 15.9 to 21.7, and BrowseComp—an evaluation of tool-augmented agent performance—jumped from 30.0 to 38.5, signaling a substantial uplift in search agent reliability. Code-related benchmarks such as LiveCodeBench held steady near 74.9, while general reasoning metrics like MMLU-Pro and GPQA-Diamond showed modest but consistent improvements. Released under an MIT license with open-source weights available on Hugging Face, the model invites developers to inspect, fine-tune, and deploy it directly, extending the openness that has become a hallmark of the DeepSeek family. SambaNova's deployment reports throughput exceeding 200 tokens per second, positioning V3.1-Terminus among the fastest open-weights reasoning models available for production agent pipelines.

Heliconedeepseek-v3.1-terminusdeepseek

Quick Info

Powered by
Provider
Helicone
Model key
deepseek-v3.1-terminus
Release date
Sep 22, 2025
Last updated
Sep 22, 2025
Knowledge cutoff
2025-09
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.27
Output token cost
$1.00

Limits

Output tokens
16,384 tokens
Context window
128,000 tokens

Latest news about DeepSeek V3.1 Terminus

Helicone

CoverageRelease Notes

The Chinese artificial intelligence startup has now released DeepSeek-V3.2-Exp, an experimental version of its current model DeepSeek-V3.1-Terminus.

Helicone

Coverage

DeepSeek-V3.1-Terminus comes just two months after the Chinese start-up launched V3.1, its most advanced model to date.

Helicone

Coverage

DeepSeek, the Chinese AI startup, has released V3.1-Terminus, a fully open source large language model designed to improve coding and search tasks while

Helicone

CoverageBenchmark

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's performance in coding and search agents. $0.27 per millio

Videos about DeepSeek V3.1 Terminus

More models around DeepSeek V3.1 Terminus