Sulat.com
AI models
Amazon Bedrock logo

Model details

gpt-oss-120b

The gpt-oss-120b is a large-scale mixture-of-experts model engineered to bring high-performance reasoning and agentic capabilities to local and self-hosted environments. With 117 billion parameters and 5.1 billion active parameters, it is designed to balance deep language understanding with the operational efficiency required for production-grade tasks. The model is specifically optimized for complex reasoning, coding, and function calling, providing developers with a versatile tool that can be integrated into agentic workflows while maintaining the flexibility of an open-weight architecture.

Built using a combination of reinforcement learning and techniques derived from frontier systems, the model benefits from a lineage that emphasizes rigorous reasoning and tool use. It is trained on the harmony response format, which is essential for its proper function and performance. The model offers configurable reasoning effort, allowing users to scale performance based on specific latency needs, and provides full access to its chain-of-thought process for easier debugging. By delivering performance that rivals proprietary systems on benchmarks like Tau-Bench and HealthBench, it serves as a robust foundation for developers looking to build, customize, and deploy AI solutions under the permissive Apache 2.0 license.

Amazon Bedrockopenai.gpt-oss-120b-1:0gpt-oss

Quick Info

Powered by
Provider
Amazon Bedrock
Model key
openai.gpt-oss-120b-1:0
Release date
Aug 5, 2025
Last updated
Aug 5, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Output tokens
16,384 tokens
Context window
128,000 tokens

Transparent token rates

Compare gpt-oss-120b pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about gpt-oss-120b

Amazon Bedrock

Official sourceAnnouncement

As AI agent deployments grow across enterprise systems, organizations need a way to bring together foundation models, tools, and governance frameworks to...

Amazon Bedrock

CoverageComparison

8x NVIDIA GB10 GPT OSS 120B Concurrency Vs TP. 8x NVIDIA GB10 Cluster Monitoring Under Load. 8x NVIDIA GB10 GPT OSS 120B Concurrency Vs TP.

Videos about gpt-oss-120b

More models around gpt-oss-120b