Sulat.com
AI models
Hugging Face logo

Model details

GPT OSS 120B

GPT OSS 120B is the larger sibling in OpenAI's gpt-oss family, released under an Apache 2.0 license that allows local or cloud hosting, modification, fine-tuning, and commercial use. The model was trained using a mix of reinforcement learning and techniques informed by OpenAI's internal frontier systems, explicitly including o3, and the launch positions it as the company's first open-weight language model release since Whisper and CLIP. Its release is accompanied by a technical paper on arXiv (identifier 2508.10925) and a Hugging Face repository that distributes the weights for community use.

Architecturally, GPT OSS 120B is built as a mixture-of-experts model with 128 experts routed per token, enabling efficient inference by activating only the relevant specialists. It is sized for data-center-grade deployment and is designed to run on a single 80 GB GPU, where OpenAI reports near-parity with o4-mini on core reasoning benchmarks and strong performance on tool use and agentic evaluations such as Tau-bench. These characteristics make the model well suited to production reasoning workloads, agentic pipelines with function calling, and fine-tuning for domain-specific applications where organizations want full control over the weights rather than relying on a hosted API.

Hugging Faceopenai/gpt-oss-120bgpt-oss

Quick Info

Powered by
Provider
Hugging Face
Model key
openai/gpt-oss-120b
Release date
Aug 5, 2025
Last updated
Aug 5, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.25
Output token cost
$0.69

Limits

Output tokens
32,768 tokens
Context window
131,072 tokens

Latest news about GPT OSS 120B

Videos about GPT OSS 120B

Recent tweets and retweets from Hugging Face

More models around GPT OSS 120B