Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
DigitalOcean logo

Model details

NVIDIA Nemotron 3 Super 120B (Public Preview)

NVIDIA Nemotron 3 Super 120B is a large hybrid-architecture language model from NVIDIA's Nemotron family, released as a public preview with openly published weights. The base variant documented on NVIDIA's Hugging Face repository uses a hybrid Latent Mixture-of-Experts design that interleaves Mamba-2 layers with MoE layers, with the repository identifier indicating BF16 precision weights. The model was trained from scratch by NVIDIA using next-token prediction and is positioned as a starting platform for further post-training such as instruction following and coding specialization, while the family maintains an official NVIDIA developer page and a published Nemotron technical report for further reference.

With open weights available and a parameter footprint of 120B (active around 12B at inference thanks to the MoE routing), the model is aimed at developers who want a self-hostable or cloud-deployed foundation model capable of reasoning and tool calling. Its training window spans the second half of 2025 through early 2026, with pre-training data current to December 2025 and post-training data through February 2026, so it reflects recent knowledge for practical assistant and agent workloads. Teams looking to combine a capable open model with a managed inference backend will find it fits scenarios where custom fine-tuning, long context, and structured outputs matter, while keeping the ability to host weights themselves on infrastructure they control.

DigitalOceannvidia-nemotron-3-super-120bnemotron

Quick Info

Powered by
Provider
DigitalOcean
Model key
nvidia-nemotron-3-super-120b
Release date
Mar 11, 2026
Last updated
Mar 11, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$0.65

Limits

Output tokens
32,768 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare NVIDIA Nemotron 3 Super 120B (Public Preview) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about NVIDIA Nemotron 3 Super 120B (Public Preview)

DigitalOcean

CoverageBenchmark

A technical write-up dated 2026-09-02 on LLMPodium examines NVIDIA Nemotron 3 Super (specifically the Nemotron-3-Super-120B-A12B variant, matching the DigitalOcean-hosted model family) and details its hybrid Mamba-Transformer mixture-of-experts topology. The architecture totals 120 billion parameters with only 12 billi The piece positions Nemotron 3 Super as targeted at agentic pipelines — multi-step coding, cyber triage, and tool-using loops — where it claims roughly 5x inference throughput and an 85.6% score on agentic PinchBench, citing the model's sparse activation and native 4-bit pre-training as the efficiency drivers. The anal

Videos about NVIDIA Nemotron 3 Super 120B (Public Preview)

More models around NVIDIA Nemotron 3 Super 120B (Public Preview)