Sulat.com
AI models
Regolo AI logo

Model details

GPT-OSS-120B

GPT-OSS-120B is a 120-billion parameter open-weight language model built on a Mixture-of-Experts architecture, activating 5.1 billion parameters per forward pass to balance capability with computational efficiency. This design enables the model to run on a single 80 GB GPU through native MXFP4 quantization, making large-scale reasoning tasks accessible without requiring massive infrastructure. Released under the flexible Apache 2.0 license, GPT-OSS-120B prioritizes transparency and customizability for developers and organizations that need to inspect, audit, or fine-tune their AI systems. The model was developed to push the frontier of open-weight reasoning models, delivering strong real-world performance at lower deployment costs compared to fully closed systems.

The training approach for GPT-OSS-120B combined reinforcement learning with techniques distilled from OpenAI's most advanced internal systems, including o3 and other frontier models. This lineage enables near-parity with o4-mini on core reasoning benchmarks while excelling at tool use, few-shot function calling, and chain-of-thought reasoning—outcomes validated on agentic evaluation suites like Tau-Bench and HealthBench, where it even surpassed proprietary predecessors. The model is compatible with the Responses API and designed for agentic workflows, supporting function calling, browsing, and structured output generation out of the box. Its availability across major cloud platforms and enterprise systems like IBM watsonx Orchestrate reflects a broader push toward bringing powerful open-weight models into regulated and high-stakes production environments.

Regolo AIgpt-oss-120bgpt-oss

Quick Info

Powered by
Provider
Regolo AI
Model key
gpt-oss-120b
Release date
Aug 5, 2025
Last updated
Aug 5, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$4.20

Limits

Output tokens
16,384 tokens
Context window
128,000 tokens

Transparent token rates

Compare GPT-OSS-120B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT-OSS-120B

Regolo AI

CoverageRelease Notes

Discover more about what's new at AWS with OpenAI GPT OSS and NVIDIA Nemotron Models Available on Amazon Bedrock in AWS GovCloud (US)

Regolo AI

Coverage

As AI agent deployments grow across enterprise systems, organizations need a way to bring together foundation models, tools, and governance frameworks to...

Videos about GPT-OSS-120B

More models around GPT-OSS-120B