Currently listed through these providers:
Model details
GPT OSS 120B
The GPT OSS 120B model represents a significant step in open-weight AI development, built on a Mixture-of-Experts architecture that incorporates 128 specialized experts with a routing system to activate relevant pathways for efficient, targeted responses. This design enables the model to deliver strong reasoning capabilities while maintaining computational efficiency. The model natively supports chain-of-thought reasoning with adjustable depth levels, allowing users to balance quality against response speed based on task demands. Native tool use capabilities including function calling, browsing, and structured output generation position it well for agentic workflows, while its Apache 2.0 licensing opens the door for local deployment, modification, fine-tuning, and commercial use without restrictive licensing constraints.
Training for this model drew on techniques informed by OpenAI's frontier systems, incorporating reinforcement learning methods alongside approaches developed through their most advanced internal models. The result is a model that achieves near-parity with proprietary systems like o4-mini on core reasoning benchmarks while running efficiently on a single 80 GB GPU, making large-scale deployment accessible to organizations without massive infrastructure overhead. On agentic evaluation suites including Tau-Bench and HealthBench, the model demonstrates strong tool use and few-shot function calling capabilities, in some cases outperforming earlier proprietary releases like o1 and GPT-4o. The combination of openWeights access, strong benchmark performance, and tool-augmented reasoning makes this model particularly well-suited for enterprise environments prioritizing data privacy, customization, and production-grade agentic applications.
Quick Info
Powered by- Provider
- Databricks
- Model key
- databricks-gpt-oss-120b
- Release date
- Aug 5, 2025
- Last updated
- Aug 5, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.072
- Output token cost
- $0.28
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Latest news about GPT OSS 120B
Videos about GPT OSS 120B
More models around GPT OSS 120B
This exact model name is also listed by 33 other providers.