Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Eden AI logo

Model details

Nemotron 3 Super 120B A12B (Nebius)

Nemotron 3 Super 120B A12B is an NVIDIA-built large language model aimed at agentic, reasoning, and conversational tasks. Its defining architectural choice is a hybrid Latent Mixture-of-Experts layout that interleaves Mamba-2 layers with MoE layers, paired with Multi-Token Prediction to accelerate generation. The result is a model that carries 120 billion total parameters but only activates 12 billion per forward pass, a sparse activation pattern that aims to balance raw capacity with inference efficiency rather than running every weight on every token.

Beyond the core architecture, the model is documented for multilingual use across English, French, German, Italian, Japanese, Spanish, and Chinese, and it is already being deployed by independent inference providers such as DeepInfra and Fireworks AI, the latter offering an NVFP4 quantized variant for on-demand serving. That third-party hosting footprint suggests practical availability across latency-sensitive and cost-sensitive production scenarios, and the multi-token prediction design is positioned to deliver faster token throughput than comparable dense checkpoints. It is a reasonable fit for teams that want strong open-weight reasoning and tool-using behavior without paying the full inference cost of a dense 120B model, especially when agentic pipelines and long-context reasoning are central requirements.

Eden AInebius/nvidia/nemotron-3-super-120b-a12bnemotron

Quick Info

Powered by
Provider
Eden AI
Model key
nebius/nvidia/nemotron-3-super-120b-a12b
Release date
Mar 11, 2026
Last updated
Mar 11, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$0.90

Limits

Output tokens
262,144 tokens
Context window
262,144 tokens

Transparent token rates

Compare Nemotron 3 Super 120B A12B (Nebius) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Nemotron 3 Super 120B A12B (Nebius)

No articles yet. Fetch the latest news to show it here.

Videos about Nemotron 3 Super 120B A12B (Nebius)

More models around Nemotron 3 Super 120B A12B (Nebius)