Currently listed through these providers:
Model details
OpenAI GPT-oss-20b
OpenAI's gpt-oss-20b is a 21-billion-parameter open-weight model with 3.6 billion active parameters, deliberately sized for lower-latency inference, on-device experimentation, and specialized deployments where the larger gpt-oss-120b would be excessive. It is part of OpenAI's gpt-oss open-models series released alongside its 120-billion-parameter sibling, and it ships under the permissive Apache 2.0 license, which removes copyleft friction for commercial, research, and fine-tuning work. Both models in the family were trained on OpenAI's harmony response format, and the smaller variant must be used with that format to function correctly, a constraint that shapes how prompts, tool calls, and structured outputs are assembled around it.
In practical terms, gpt-oss-20b is positioned as a developer-friendly reasoning and agentic model: configurable reasoning effort across low, medium, and high settings lets teams trade latency for depth on a per-task basis, while native function calling and full chain-of-thought access support debugging and trust in outputs. Independent technical coverage of the gpt-oss family highlights post-training work aimed at reasoning, tool use, and agentic behavior, alongside benchmarks and comparisons against other open and proprietary systems that contextualize where the 20B variant sits. That combination of open weights, an active-parameter footprint suited to single-accelerator or local serving, and harmony-format agentic features makes gpt-oss-20b a fit for teams wanting OpenAI-style reasoning behavior with the control and customization that an open model affords.
Quick Info
Powered by- Provider
- DigitalOcean
- Model key
- openai-gpt-oss-20b
- Release date
- Aug 5, 2025
- Last updated
- Apr 16, 2026
- Knowledge cutoff
- 2024-06
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.05
- Output token cost
- $0.45
Limits
- Output tokens
- 4,096 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare OpenAI GPT-oss-20b pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about OpenAI GPT-oss-20b
No articles yet. Fetch the latest news to show it here.