Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

Nemotron 3.5 Content Safety

Nemotron 3.5 Content Safety is a small language model published by NVIDIA, packaged as a NIM container on NVIDIA NGC and built on Google's Gemma-3-4B-it base, which NVIDIA then fine-tuned on multimodal, multilingual, and reasoning-oriented content-safety datasets. The model is positioned as a compact 4B-parameter safety classifier that extends the earlier Nemotron 3 Content Safety, adding coverage for prompts, responses, and images so a single call can judge both text and visual content. Because it is released as open weights, teams can self-host it on NVIDIA-accelerated infrastructure and integrate it with common inference frameworks for real-time moderation in production pipelines.

In practice, the model is intended for developers and enterprises that need to moderate AI inputs and outputs across languages and modalities while staying inside their own governance rules. It supports a 23-category safety taxonomy, customizable policy reasoning with concise traces before each verdict, and multilingual moderation for a dozen languages out of the box, making it suitable for global deployments. An independent guardrail benchmark run by Artificial Analysis in partnership with NVIDIA placed Nemotron 3.5 Content Safety among the specialist safety classifiers evaluated for F1 score, recall, specificity, and end-to-end latency on open datasets like WildGuardTest, ToxicChat, and XSTest, giving adopters a public reference point for its balance of catching unsafe content without over-refusing safe prompts.

Kilo Gatewaynvidia/nemotron-3.5-content-safetynemotron

Quick Info

Powered by
Provider
Kilo Gateway
Model key
nvidia/nemotron-3.5-content-safety
Release date
Jun 4, 2026
Last updated
Jun 4, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.20
Output token cost
$0.20

Limits

Output tokens
117,964 tokens
Context window
131,072 tokens

Transparent token rates

Compare Nemotron 3.5 Content Safety pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Nemotron 3.5 Content Safety

No articles yet. Fetch the latest news to show it here.

Videos about Nemotron 3.5 Content Safety

More models around Nemotron 3.5 Content Safety