Currently listed through these providers:
Model details
OpenAI GPT-OSS 20b
The GPT-OSS 20b is designed as a compact, high-performance model within the gpt-oss family, specifically engineered for lower latency and specialized deployment scenarios. With 21 billion total parameters and 3.6 billion active parameters, it provides a balanced architecture that excels in agentic tasks, such as function calling, while maintaining a smaller footprint than its larger counterparts. The model is built to support configurable reasoning efforts, allowing users to adjust performance based on specific latency requirements, and provides full access to its chain-of-thought process to improve transparency and debugging during complex workflows.
Trained on the harmony response format, this model requires adherence to this specific structure to ensure correct output generation. Its release under the permissive Apache 2.0 license encourages broad experimentation and commercial integration, making it a practical choice for developers looking to fine-tune models for custom use cases. The model is optimized for local execution, with demonstrated compatibility across various hardware platforms, including support for Ryzen and Radeon systems. By offering a blend of agentic utility and local accessibility, it serves as a flexible foundation for building and deploying specialized AI applications.
Quick Info
Powered by- Provider
- Helicone
- Model key
- gpt-oss-20b
- Release date
- Jun 1, 2024
- Last updated
- Jun 1, 2024
- Knowledge cutoff
- 2024-06
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.05
- Output token cost
- $0.20
Limits
- Output tokens
- 131,072 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare OpenAI GPT-OSS 20b pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about OpenAI GPT-OSS 20b
No articles yet. Fetch the latest news to show it here.