Sulat.com
AI models
Groq logo

Model details

GPT OSS 120B

The GPT-OSS 120B represents OpenAI's first major venture into fully open-weight language models since releasing Whisper and CLIP, offering Apache 2.0 licensed access to a model designed for powerful reasoning, agentic tasks, and versatile developer use cases. The architecture is notable for its ability to fit into a single 80GB GPU like an NVIDIA H100 or AMD MI300X, with the model supporting configurable reasoning effort levels so developers can balance depth of analysis against latency needs. Built-in capabilities include full chain-of-thought visibility for debugging and trust verification, native function calling for agentic workflows, and full parameter fine-tuning support for customization. The harmony response format underpins the model's interaction paradigm, enabling structured outputs that integrate well with enterprise deployment patterns.

Training and post-training decisions reflect a focus on practical enterprise and developer fit rather than purely academic benchmarks. The harmony response format mentioned in model documentation suggests deliberate post-training to shape how the model structures its outputs and reasoning traces. Enterprise adoption is already materializing through integrations like IBM watsonx Orchestrate on Amazon Bedrock, indicating the model is being evaluated for production AI agent deployments. The open-weight approach under Apache 2.0 licensing removes commercial barriers, while the existence of a 50% compressed derivative variant demonstrates an active ecosystem building around the base model. For developers and organizations, this positioning suggests a model that balances the reasoning and agentic capabilities of larger systems with deployment flexibility that fits real-world infrastructure constraints.

Groqopenai/gpt-oss-120bgpt-oss

Quick Info

Powered by
Provider
Groq
Model key
openai/gpt-oss-120b
Release date
Aug 5, 2025
Last updated
Oct 21, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Output tokens
65,536 tokens
Context window
131,072 tokens

Latest news about GPT OSS 120B

Groq

CoverageComparison

8x NVIDIA GB10 GPT OSS 120B Concurrency Vs TP. 8x NVIDIA GB10 Cluster Monitoring Under Load. 8x NVIDIA GB10 GPT OSS 120B Concurrency Vs TP.

Groq

CoverageRelease Notes

OpenAI is going back to its open-source roots. Sort of.

Groq

Coverage

As AI agent deployments grow across enterprise systems, organizations need a way to bring together foundation models, tools, and governance frameworks to...

Groq

CoverageRelease Notes

Multiverse Computing releases a compressed version of OpenAI's gpt-oss-120B

Groq

Coverage

HyperNova 60B 2602, a 50% compressed version of OpenAI’s gpt-oss-120B, accelerates Multiverse’s plans to deliver hyper-efficient, high-performance models for free to developersDONOSTIA, Spain, Feb. 24, 2026 (GLOBE NEWSWIRE) -- Multiverse Computing, the leader in AI model compression, today announced the release of Hype

Groq

Coverage

HyperNova 60B 2602, a 50% compressed version of OpenAI’s gpt-oss-120B, accelerates Multiverse’s plans to deliver hyper-efficient, high-performance models for free to developersDONOSTIA, Spain, Feb. 24, 2026 (GLOBE NEWSWIRE) -- Multiverse Computing, the leader in AI model compression, today announced the release of Hype

Videos about GPT OSS 120B

More models around GPT OSS 120B