Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Clarifai logo

Model details

GPT OSS 20B

GPT OSS 20B sits at the lighter end of OpenAI's gpt-oss family, released as an open-weight model under the permissive Apache 2.0 license alongside its larger sibling. The model carries 21B total parameters with 3.6B active parameters, a sparse configuration that lets it target lower-latency, on-device, or specialized deployments where running a frontier-scale model is impractical. Like its sibling, it was trained specifically on OpenAI's harmony response format, so downstream systems must wrap prompts and completions in that format to get correct behavior; using a generic chat template will degrade output quality.

For practical work, the model is designed to expose its reasoning process end-to-end, with configurable reasoning effort across low, medium, and high settings so developers can trade latency against depth of thought on a per-request basis. It supports parameter fine-tuning and ships with native function-calling for agentic pipelines, making it a flexible base for building custom assistants, retrieval-augmented systems, or domain-specific copilots that need to run locally. The combination of permissive licensing, adjustable reasoning, and an accessible parameter footprint positions it well for experimentation, private deployments, and cost-sensitive production use where full-size frontier models would be overkill.

Clarifaiopenai/chat-completion/models/gpt-oss-20bgpt-oss

Quick Info

Powered by
Provider
Clarifai
Model key
openai/chat-completion/models/gpt-oss-20b
Release date
Aug 5, 2025
Last updated
Dec 12, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.045
Output token cost
$0.18

Limits

Output tokens
16,384 tokens
Context window
131,072 tokens

Transparent token rates

Compare GPT OSS 20B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT OSS 20B

Clarifai

CoverageBenchmark

BenchLM's tracking page for GPT-OSS 20B (data as of September 2, 2026) records the model as an open-weight, non-reasoning-category release from August 5, 2025 with a 128K context window. The aggregate capability score sits at 42.4/100, ranking the model 188 of 230 in the catalog, with its strongest eligible category be The capability radar shows only Coding as a rank-eligible axis (13th percentile within its cohort), while Agentic, Reasoning, Knowledge, Math, Multilingual, Multimodal, and Instruction-Following are listed as not eligible/not ranked. Reported throughput is 313 tok/s with a 0.65s first-token latency against field median

Videos about GPT OSS 20B

More models around GPT OSS 20B