Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Amazon Bedrock logo

Model details

NVIDIA Nemotron Nano 3 30B

NVIDIA Nemotron Nano 3 30B represents the third generation of NVIDIA's Nemotron Nano family, built with a hybrid MoE architecture that blends Mamba's efficient sequential processing with the Transformer's powerful attention mechanisms. This architectural fusion gives the model a natural edge in agentic production workloads, where rapid reasoning across long, multi-step interactions matters most. With 30 billion parameters and a design that prioritizes inference efficiency, the model targets developers and enterprises seeking capable, responsive AI that fits comfortably into real-world deployment budgets.

As an officially open-weights model, Nemotron Nano 3 invites scrutiny and customization, reflecting NVIDIA's commitment to transparency in the agentic AI era. The architecture has demonstrated up to 13x faster token generation in supported deployments, a metric that speaks directly to its suitability for latency-sensitive applications. FriendliAI's early adoption as a launch partner underscores the model's appeal for high-throughput, cost-conscious inference at scale across diverse cloud and on-premise environments.

Amazon Bedrocknvidia.nemotron-nano-3-30bnemotron

Quick Info

Powered by
Provider
Amazon Bedrock
Model key
nvidia.nemotron-nano-3-30b
Release date
Dec 15, 2025
Last updated
Dec 23, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.06
Output token cost
$0.24

Limits

Output tokens
8,192 tokens
Context window
262,144 tokens

Latest news about NVIDIA Nemotron Nano 3 30B

No articles yet. Fetch the latest news to show it here.

Videos about NVIDIA Nemotron Nano 3 30B

More models around NVIDIA Nemotron Nano 3 30B