Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Clarifai logo

Model details

GPT OSS 120B High Throughput

GPT OSS 120B High Throughput is listed as a member of the gpt-oss family, a line of openly released OpenAI language models that organizations can self-host or access through third-party inference providers. On Clarifai's catalog it appears under the canonical model key openai/chat-completion/models/gpt-oss-120b-high-throughput, with Clarifai's own inference documentation linked from the provider page to explain how requests are routed and billed. The "high throughput" designation within this model key points to a serving profile tuned for batch and parallel workloads rather than a separate model variant, making it a practical option when many concurrent completions need to be processed efficiently. Because the weights are openly available, this 120B-parameter model is positioned for teams that want strong reasoning and tool-use behavior without depending on a closed API, and the Clarifai listing surfaces it as a turnkey hosted path for that deployment. Its long context window allows it to ingest substantial code bases, documents, or multi-turn agent traces in a single request, while the throughput-oriented configuration makes it well suited to pipelines such as bulk classification, evaluation harnesses, and agent backends that issue many calls in parallel. Practitioners evaluating open-weight frontier-tier reasoning models will find this Clarifai-hosted gpt-oss 120B option a reasonable balance of capability, openness, and serving scale.

The gpt-oss family reflects OpenAI's broader move toward publishing competitive open-weight checkpoints that can match closed models on reasoning-heavy tasks while remaining inspectable and adaptable. Hosting such a large open model at high throughput requires meaningful infrastructure, and Clarifai's inclusion of the 120B variant alongside its inference docs signals an intent to serve enterprise and developer workloads that need both capability and scale. Compared with smaller open models, a 120B-parameter checkpoint generally delivers stronger multi-step reasoning, more reliable tool calling, and better handling of long structured prompts, which are common needs for production assistants and automation agents. Choosing this hosted configuration is most attractive for teams that want frontier-tier open-model behavior without standing up their own GPU clusters, and who value the flexibility of open weights should they later move inference in-house.

Clarifaiopenai/chat-completion/models/gpt-oss-120b-high-throughputgpt-oss

Quick Info

Powered by
Provider
Clarifai
Model key
openai/chat-completion/models/gpt-oss-120b-high-throughput
Release date
Aug 5, 2025
Last updated
Feb 25, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.09
Output token cost
$0.36

Limits

Output tokens
16,384 tokens
Context window
131,072 tokens

Latest news about GPT OSS 120B High Throughput

No articles yet. Fetch the latest news to show it here.

Videos about GPT OSS 120B High Throughput

More models around GPT OSS 120B High Throughput