Sulat.com
AI models
ai& logo

Model details

GPT OSS 120B

GPT OSS 120B is OpenAI's open-weight large language model built around a Mixture-of-Experts design, with 117 billion total parameters of which 5.1 billion activate per forward pass, allowing it to run efficiently on a single high-end GPU using native MXFP4 quantization. It is distributed under the Apache 2.0 license alongside its smaller sibling, the 20B variant, giving developers full access to the weights. The model was trained with a combination of reinforcement learning and techniques drawn from OpenAI's frontier research lineage, including o3, and is positioned as a flexible foundation for high-reasoning, agentic, and general-purpose production workloads.

In practice, GPT OSS 120B is engineered for agentic workflows, offering configurable reasoning depth, full chain-of-thought access, and native tool use including function calling, browsing, and structured output generation. It achieves near-parity with OpenAI's o4-mini on core reasoning benchmarks and posts strong results on agentic evaluations such as Tau-Bench and HealthBench, where it can outperform some proprietary models. The Responses API compatibility and broad ecosystem support make it well suited for teams building multi-step agents, code-execution pipelines, or research applications that need transparent reasoning traces without the constraints of a closed-source deployment.

ai&openai/gpt-oss-120bgpt-oss

Quick Info

Powered by
Provider
ai&
Model key
openai/gpt-oss-120b
Release date
Aug 5, 2025
Last updated
Aug 5, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Output tokens
32,768 tokens
Context window
131,072 tokens

Transparent token rates

Compare GPT OSS 120B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT OSS 120B

ai&

Coverage

The pi-aiand Pi extension, version 1.0.1 and published July 28, 2026, adds ai& as a provider for eight open-weight models. Its package description lists openai/gpt-oss-120b with a 131K-token context window and says the extension can add the models to Pi’s /model command without a separate provider block or base URL con For ai&’s organization-scoped catalog, the page lists GPT-OSS 120B at ¥25 per million input tokens (approximately $0.16) and ¥95 per million output tokens (approximately $0.59). It says the bundled catalog is synchronized from an authenticated GET /v1/models response, so the displayed model availability and prices may

ai&

CoverageBenchmark

OpenRouter describes OpenAI’s gpt-oss-120b as a 117B-parameter mixture-of-experts model that activates 5.1B parameters per forward pass and is optimized for a single H100 using native MXFP4 quantization. The page lists a 131K context window, configurable reasoning, chain-of-thought access, and native function calling, The model was released on August 5, 2025, and the page gives a standard displayed price of $0.03 per million input tokens and $0.17 per million output tokens. It also provides a dated cross-provider matrix with different latency, throughput, uptime, and pricing figures, but the excerpt does not identify ai& as a provid

Videos about GPT OSS 120B

More models around GPT OSS 120B