Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Eden AI logo

Model details

GPT OSS Safeguard 20B (Groq)

GPT OSS Safeguard 20B is an open-weight reasoning model post-trained from the underlying GPT-OSS family and released by OpenAI under the Apache 2.0 license alongside the GPT-OSS usage policy. It is text-only and was designed specifically to reason from a provided policy in order to label content against that policy, rather than to serve as a general-purpose chat assistant. The technical report positions the larger sibling and this 20B variant as customizable classifiers that expose full chain-of-thought reasoning, support configurable reasoning effort levels (low, medium, and high), and integrate with Structured Outputs for downstream pipelines.

On the policy-classification benchmarks reported by OpenAI, the model shows the qualitative strengths its creators highlight: it outperforms GPT-5-thinking on multi-policy accuracy and slightly edges out OpenAI's internal Safety Reasoner on the 2022 OpenAI Moderation evaluation, while remaining a compact 20B option that the report describes as preferable for moderation tasks. Multilingual performance on MMMLU is reported to track closely with the base GPT-OSS models across all reasoning effort levels, indicating that the safety-focused post-training did not regress language coverage. Practical fit centers on internal moderation and content-labeling systems where a deployable open-weight reasoner that can follow a supplied rubric is more valuable than a broad conversational model.

Eden AIgroq/openai/gpt-oss-safeguard-20bgpt-oss

Quick Info

Powered by
Provider
Eden AI
Model key
groq/openai/gpt-oss-safeguard-20b
Release date
Oct 29, 2025
Last updated
Oct 29, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.075
Output token cost
$0.30

Limits

Output tokens
131,072 tokens
Context window
131,072 tokens

Transparent token rates

Compare GPT OSS Safeguard 20B (Groq) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT OSS Safeguard 20B (Groq)

Eden AI

CoverageBenchmark

OpenAI's hosted PDF of the technical report, dated October 29, 2025, provides the canonical first-party documentation for gpt-oss-safeguard-20b (and the 120b sibling). It explicitly names both variants as open-weight reasoning models post-trained from the gpt-oss models and trained to reason from a provided policy for The PDF reports that the safeguard models are fine-tunes of their gpt-oss counterparts and were trained without additional biological or cybersecurity data, so prior worst-case scenario estimates from the gpt-oss release are cross-applied. Beyond the introduction, the report's sections cover safety classification perfo

Videos about GPT OSS Safeguard 20B (Groq)

More models around GPT OSS Safeguard 20B (Groq)