Sulat.com
AI models
Nebius Token Factory logo

Model details

gpt-oss-120b

Built on a scalable transformer architecture, gpt-oss-120b is designed to handle complex, multi-step reasoning tasks while maintaining high efficiency. With 117 billion parameters and 5.1 billion active parameters, the model is optimized to run on a single 80 GB GPU, making it a practical choice for organizations looking to integrate powerful AI into their own infrastructure. It is specifically engineered for agentic workflows, offering robust support for function calling, structured outputs, and retrieval integration, which helps developers build durable systems with minimal additional code.

The model benefits from a training lineage that incorporates reinforcement learning and techniques derived from frontier systems. It utilizes a specialized harmony response format and variable effort reasoning training, allowing users to adjust the model's reasoning intensity based on specific latency and performance needs. By providing access to its chain-of-thought process, the model enables greater transparency and easier debugging for developers. Released under the Apache 2.0 license, it serves as a versatile tool for those seeking to customize and deploy state-of-the-art reasoning capabilities in diverse, real-world environments.

Nebius Token Factoryopenai/gpt-oss-120b

Quick Info

Powered by
Provider
Nebius Token Factory
Model key
openai/gpt-oss-120b
Release date
Jan 10, 2026
Last updated
Feb 4, 2026
Knowledge cutoff
2025-09
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Input tokens
124,000 tokens
Output tokens
8,192 tokens
Context window
128,000 tokens

Latest news about gpt-oss-120b

Nebius Token Factory

Coverage

As AI agent deployments grow across enterprise systems, organizations need a way to bring together foundation models, tools, and governance frameworks to...

Nebius Token Factory

Coverage

HyperNova 60B 2602, a 50% compressed version of OpenAI’s gpt-oss-120B, accelerates Multiverse’s plans to deliver hyper-efficient, high-performance models for free to developersDONOSTIA, Spain, Feb. 24, 2026 (GLOBE NEWSWIRE) -- Multiverse Computing, the leader in AI model compression, today announced the release of Hype

Videos about gpt-oss-120b