Currently listed through these providers:
Model details
NVIDIA Nemotron Nano 3 30B
NVIDIA Nemotron Nano 3 30B represents the third generation of NVIDIA's Nemotron Nano family, built with a hybrid MoE architecture that blends Mamba's efficient sequential processing with the Transformer's powerful attention mechanisms. This architectural fusion gives the model a natural edge in agentic production workloads, where rapid reasoning across long, multi-step interactions matters most. With 30 billion parameters and a design that prioritizes inference efficiency, the model targets developers and enterprises seeking capable, responsive AI that fits comfortably into real-world deployment budgets.
As an officially open-weights model, Nemotron Nano 3 invites scrutiny and customization, reflecting NVIDIA's commitment to transparency in the agentic AI era. The architecture has demonstrated up to 13x faster token generation in supported deployments, a metric that speaks directly to its suitability for latency-sensitive applications. FriendliAI's early adoption as a launch partner underscores the model's appeal for high-throughput, cost-conscious inference at scale across diverse cloud and on-premise environments.
Quick Info
Powered by- Provider
- Amazon Bedrock
- Model key
- nvidia.nemotron-nano-3-30b
- Release date
- Dec 15, 2025
- Last updated
- Dec 23, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.06
- Output token cost
- $0.24
Limits
- Output tokens
- 8,192 tokens
- Context window
- 262,144 tokens
Latest news about NVIDIA Nemotron Nano 3 30B
No articles yet. Fetch the latest news to show it here.