Currently listed through these providers:
Model details
GPT OSS Safeguard 120B
As an open-weight safety reasoning model in the gpt-oss family, this release is positioned as a fine-tuned descendant of OpenAI's earlier gpt-oss open models, distributed under the same permissive Apache 2.0 license. A 120B-parameter variant (reported as roughly 117B total with around 5.1B active parameters in a mixture-of-experts configuration) is paired with a smaller 20B sibling, and both are available for download from the OpenAI organization on Hugging Face. Hosting providers such as Fireworks AI additionally expose the weights for on-demand deployment, broadening access beyond direct self-hosting.
The defining design choice is that the model is built to interpret a developer-supplied safety policy directly at inference time rather than relying on a fixed, trained-in classifier. It can label user messages, completions, or full chats, and it returns chain-of-thought reasoning that developers can audit to understand each decision. Because the policy is injected per request, teams can iterate on rule wording, expand coverage, or swap policies for new product contexts without retraining, which makes the model a flexible fit for organizations that want transparent, customizable content moderation rather than a black-box safety filter.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- openai/gpt-oss-safeguard-120b
- Release date
- Oct 29, 2025
- Last updated
- Oct 29, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.15
- Output token cost
- $0.60
Limits
- Input tokens
- 112,000 tokens
- Output tokens
- 16,000 tokens
- Context window
- 128,000 tokens
Latest news about GPT OSS Safeguard 120B
Videos about GPT OSS Safeguard 120B
Recent tweets and retweets from Vercel AI Gateway
More models around GPT OSS Safeguard 120B
This exact model name is also listed by 4 other providers.