Currently listed through these providers:
Model details
GPT OSS 120B
GPT OSS 120B is the larger sibling in OpenAI's gpt-oss family, released under an Apache 2.0 license that allows local or cloud hosting, modification, fine-tuning, and commercial use. The model was trained using a mix of reinforcement learning and techniques informed by OpenAI's internal frontier systems, explicitly including o3, and the launch positions it as the company's first open-weight language model release since Whisper and CLIP. Its release is accompanied by a technical paper on arXiv (identifier 2508.10925) and a Hugging Face repository that distributes the weights for community use.
Architecturally, GPT OSS 120B is built as a mixture-of-experts model with 128 experts routed per token, enabling efficient inference by activating only the relevant specialists. It is sized for data-center-grade deployment and is designed to run on a single 80 GB GPU, where OpenAI reports near-parity with o4-mini on core reasoning benchmarks and strong performance on tool use and agentic evaluations such as Tau-bench. These characteristics make the model well suited to production reasoning workloads, agentic pipelines with function calling, and fine-tuning for domain-specific applications where organizations want full control over the weights rather than relying on a hosted API.
Quick Info
Powered by- Provider
- Hugging Face
- Model key
- openai/gpt-oss-120b
- Release date
- Aug 5, 2025
- Last updated
- Aug 5, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.25
- Output token cost
- $0.69
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Latest news about GPT OSS 120B
Videos about GPT OSS 120B
Recent tweets and retweets from Hugging Face
More models around GPT OSS 120B
This exact model name is also listed by 33 other providers.