Model details
gpt-oss-20b
GPT-OSS-20B marks a deliberate shift in OpenAI's strategy, representing the company's first fully open-source foundation model release since Whisper and CLIP. Built with a Mixture-of-Experts architecture, the model activates 3.6 billion parameters per forward pass while maintaining a compact 21 billion parameter footprint overall. This design choice prioritizes accessibility over raw scale—the architecture enables meaningful reasoning and tool-use capabilities while remaining deployable on consumer-grade or single-GPU cloud hardware. Released under the Apache 2.0 license, the model can be run locally, modified, fine-tuned, and commercialized without restriction.
The model arrives configured with OpenAI's Harmony response format and includes native support for configurable reasoning levels, function calling, and structured outputs—capabilities that position it for agentic workflows and multi-step task completion. Its combination of open weights, reasoning capability, and tool interaction makes it particularly suited for developers who need to inspect, customize, or self-host powerful language model infrastructure without vendor lock-in. The architecture's emphasis on efficiency over parameter count reflects a design philosophy tuned for latency-sensitive applications and cost-effective deployment at scale.
Quick Info
Powered by- Provider
- OVHcloud AI Endpoints
- Model key
- gpt-oss-20b
- Release date
- Aug 28, 2025
- Last updated
- Aug 28, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.05
- Output token cost
- $0.18
Limits
- Output tokens
- 131,072 tokens
- Context window
- 131,072 tokens
Latest news about gpt-oss-20b
No articles yet. Fetch the latest news to show it here.