Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

OpenAI GPT OSS 120B

GPT OSS 120B is built as a Mixture-of-Experts model with SwiGLU activations and learned attention sinks baked into its architecture, targeting the sweet spot between raw capability and practical deployment. The design philosophy prioritizes production-grade reasoning workloads that can run on a single high-end GPU, which means balancing massive total parameter counts against manageable active parameter counts during inference. As a reasoning-focused model, it natively supports chain-of-thought processing and lets developers dial in reasoning effort based on latency or quality needs, making it adaptable for everything from quick-turnaround tasks to complex multi-step agentic pipelines. The Apache 2.0 licensing removes barriers for commercial and enterprise use, positioning this as a genuinely open alternative for organizations that previously had limited options for on-premises high-reasoning models.

The model was trained specifically on OpenAI's harmony response format, which shapes how instructions and outputs are structured through the pipeline. This explicit training format is why the documentation emphasizes using the harmony format consistently. Beyond the base training, the model exposes fine-tuning pathways so teams can specialize it for domain-specific tasks without losing the core reasoning foundation. Agentic capabilities come through native function calling, structured output generation, and tool integration, making it suitable for building autonomous pipelines that handle multi-step workflows. The MoE architecture with its selective parameter activation means you get substantial reasoning power while keeping inference costs and memory footprint predictable, which is a practical advantage for teams deploying at scale.

NovitaAIopenai/gpt-oss-120b

Quick Info

Powered by
Provider
NovitaAI
Model key
openai/gpt-oss-120b
Release date
Aug 6, 2025
Last updated
Aug 6, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.05
Output token cost
$0.25

Limits

Output tokens
32,768 tokens
Context window
131,072 tokens

Latest news about OpenAI GPT OSS 120B

NovitaAI

CoverageRelease Notes

/PRNewswire/ -- As demand for open-source AI infrastructure grows, Novita AI is establishing itself as the inference provider for developers and engineering..., /PRNewswire/ -- As demand for open-source AI infrastructure grows, Novita AI is establishing itself as the inference provider for developers and engineering...

Videos about OpenAI GPT OSS 120B