Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenCode Zen logo

Model details

Nemotron 3 Ultra Free

Nemotron 3 Ultra Free is an open-weights large language model developed by NVIDIA and positioned for enterprise AI, advanced reasoning, coding, and agentic workflows that benefit from a very large parameter count and an extended context window. Independent comparison coverage characterizes the underlying Nemotron 3 Ultra as an NVIDIA ultra-scale architecture in the roughly 340 billion parameter range, with the exact size treated as proprietary, and frames the model as suitable for long-context, reasoning-heavy, and tool-driven tasks. The Free designation reflects zero-cost access through hosted channels rather than a change to the underlying weights, letting teams evaluate the model without separate licensing friction.

Because the weights are openly distributed, practitioners can self-host, fine-tune, or audit the model for sensitive workloads, while still taking advantage of NVIDIA's focus on agentic behavior and tool use that fits multi-step automation, retrieval-augmented pipelines, and code generation. The combination of a very long context window and NVIDIA's reasoning-oriented training emphasis makes it well suited to complex document analysis, codebase reasoning, and orchestrated assistants that need to retain large amounts of information in a single session. It is a strong fit for organizations that want an open-weights foundation model from a major hardware-aligned AI lab and need reasoning, coding, and agent-style capabilities in a single system.

OpenCode Zennemotron-3-ultra-freenemotron-free

Quick Info

Powered by
Provider
OpenCode Zen
Model key
nemotron-3-ultra-free
Release date
Jun 4, 2026
Last updated
Jun 4, 2026
Knowledge cutoff
2026-02
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
128,000 tokens
Context window
1,000,000 tokens

Latest news about Nemotron 3 Ultra Free

OpenCode Zen

CoverageBenchmark

OpenRouter's free-tier listing for NVIDIA Nemotron 3 Ultra (slug nvidia/nemotron-3-ultra-550b-a55b:free) confirms the model is a hybrid Transformer-Mamba mixture-of-experts with 550B total parameters and 55B active per token, text in/out modalities, a 1M-token context window, and a release date of June 4, 2026. It is p For the Free variant, price is listed as Free, and the NVIDIA Free provider shows a P50 throughput of about 8 tokens/second, P50 latency of 28.05 seconds, and 98.30% uptime over the prior week. The page carries an explicit notice that the free endpoint logs session data for security and product improvement under NVIDIA

Videos about Nemotron 3 Ultra Free

More models around Nemotron 3 Ultra Free