Nebius Token Factory
As AI agent deployments grow across enterprise systems, organizations need a way to bring together foundation models, tools, and governance frameworks to...
Model details
Built on a scalable transformer architecture, gpt-oss-120b is designed to handle complex, multi-step reasoning tasks while maintaining high efficiency. With 117 billion parameters and 5.1 billion active parameters, the model is optimized to run on a single 80 GB GPU, making it a practical choice for organizations looking to integrate powerful AI into their own infrastructure. It is specifically engineered for agentic workflows, offering robust support for function calling, structured outputs, and retrieval integration, which helps developers build durable systems with minimal additional code.
The model benefits from a training lineage that incorporates reinforcement learning and techniques derived from frontier systems. It utilizes a specialized harmony response format and variable effort reasoning training, allowing users to adjust the model's reasoning intensity based on specific latency and performance needs. By providing access to its chain-of-thought process, the model enables greater transparency and easier debugging for developers. Released under the Apache 2.0 license, it serves as a versatile tool for those seeking to customize and deploy state-of-the-art reasoning capabilities in diverse, real-world environments.
Nebius Token Factory
As AI agent deployments grow across enterprise systems, organizations need a way to bring together foundation models, tools, and governance frameworks to...
Nebius Token Factory
HyperNova 60B 2602, a 50% compressed version of OpenAI’s gpt-oss-120B, accelerates Multiverse’s plans to deliver hyper-efficient, high-performance models for free to developersDONOSTIA, Spain, Feb. 24, 2026 (GLOBE NEWSWIRE) -- Multiverse Computing, the leader in AI model compression, today announced the release of Hype