Sulat.com
AI models
Amazon Bedrock logo

Model details

GPT OSS Safeguard 120B

The GPT OSS Safeguard 120B is a safety reasoning model that represents a shift away from traditional content classifiers toward policy-grounded judgment. Rather than training fixed decision boundaries from labeled examples, this model interprets a developer-supplied safety policy at inference time, applying it to classify user messages, completions, and full conversations. It employs chain-of-thought reasoning that developers can inspect to understand how conclusions are reached, and it supports configurable reasoning effort levels so teams can trade off depth against latency depending on the use case. This design gives developers explicit control over where the policy lines fall, enabling rapid iteration as safety requirements evolve.

The model is a post-trained derivative of the open-weight GPT OSS family, developed with input from the open-source community and released under a permissive Apache 2.0 license. It was initially built for internal use at OpenAI before becoming publicly available, and it pairs with the Responses API for deployment. The 120B variant is optimized for production workloads and can run on a single H100 GPU, making it accessible for teams without massive infrastructure. Unlike models designed for end-user interaction, this model is purpose-built to serve as a safety classification layer behind the scenes, classifying content with justified decisions and supporting structured outputs for integration into broader moderation pipelines.

Amazon Bedrockopenai.gpt-oss-safeguard-120bgpt-oss

Quick Info

Powered by
Provider
Amazon Bedrock
Model key
openai.gpt-oss-safeguard-120b
Release date
Oct 29, 2025
Last updated
Oct 29, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Output tokens
16,384 tokens
Context window
128,000 tokens

Transparent token rates

Compare gpt-oss pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT OSS Safeguard 120B

Amazon Bedrock

Coverage

OpenAI has launched gpt-oss-safeguard, open-weight AI safety reasoning models hosted on Hugging Face. The models allow developers to apply custom content policies, improving transparency and flexibility in moderation. Built with partner ROOST, they aim to enhance responsible AI development. Read the, OpenAI has launche

Amazon Bedrock

Coverage

OpenAI on Wednesday announced two new artificial intelligence reasoning models — gpt-oss-safeguard-120b and gpt-oss-safeguard-20b — designed...

Amazon Bedrock

Coverage

The news blog specialized in Japanese culture, odd news, gadgets and all other funny stuffs. Updated everyday.

Amazon Bedrock

Coverage

OpenAI introduces gpt-oss-safeguard—open-weight reasoning models for safety classification that let developers apply and iterate on custom policies.

Videos about GPT OSS Safeguard 120B

More models around GPT OSS Safeguard 120B