Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Nemotron 3 Nano 30B A3B

Nemotron 3 Nano 30B A3B is a large language model developed and trained from scratch by NVIDIA as a single, unified system that can handle both reasoning and non-reasoning tasks. Rather than relying on separate models for different jobs, it first produces a reasoning trace and then delivers a final answer, and users can toggle that reasoning behavior through a flag in the chat template. Disabling the trace gives a leaner response style, while leaving it on typically produces stronger results on harder prompts that genuinely benefit from step-by-step thinking, making the model adaptable across everyday chat and more analytical workloads.

The architecture is what gives the model its efficiency profile. It uses a hybrid Mixture-of-Experts design that mixes 23 Mamba-2 and MoE layers with 6 attention layers, where each MoE layer contains 128 routed experts plus 1 shared expert and activates 6 experts per token. That structure yields roughly 3.5B active parameters out of 30B total, so it behaves like a much smaller model at inference time while retaining the capacity of a larger one. The generous context window makes it well suited to long-document summarization, code repositories, multi-turn analytical sessions, and any task where maintaining coherence across very large inputs matters.

Vercel AI Gatewaynvidia/nemotron-3-nano-30b-a3bnemotron

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
nvidia/nemotron-3-nano-30b-a3b
Release date
Dec 15, 2025
Last updated
Dec 15, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.05
Output token cost
$0.20

Limits

Output tokens
262,144 tokens
Context window
262,144 tokens

Transparent token rates

Compare Nemotron 3 Nano 30B A3B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Nemotron 3 Nano 30B A3B

OpenRouter

Coverage

HokAI's model hub entry explicitly names the NVIDIA Nemotron 3 Nano 30B-A3B and describes it as an open-weight language model built by NVIDIA, released December 14, 2025 as the entry tier of the Nemotron 3 family. It reports benchmark scores from NVIDIA's technical report (arXiv 2512.20848): 73.04% on GPQA, 68.25% on L For tool-use workloads, HokAI reports that giving Nemotron 3 Nano tools lifts AIME accuracy from 89.06% to 99.17%, positioning it as a strong fit for agentic tool-calling pipelines. The page also notes an Artificial Analysis score of 38.8% on SWE-bench Verified. The license is listed as the NVIDIA Open Model License wi

Videos about Nemotron 3 Nano 30B A3B

More models around Nemotron 3 Nano 30B A3B