Sulat.com
AI models
SiliconFlow logo

Model details

openai/gpt-oss-20b

The smaller model in the gpt-oss family is built for lower-latency, local, and specialized workloads. It has 21 billion total parameters with 3.6 billion active at a time, and the series was trained around OpenAI’s Harmony response format, which the model card says is required for correct operation.

This model is intended for reasoning, tool-oriented work, and agentic applications. Its reasoning effort can be set to low, medium, or high to balance output depth and response time, while parameter fine-tuning supports customization. The Apache 2.0 license also makes it suitable for experimentation, commercial deployment, and other projects that benefit from an adaptable, locally runnable foundation.

SiliconFlowopenai/gpt-oss-20bgpt-oss

Quick Info

Powered by
Provider
SiliconFlow
Model key
openai/gpt-oss-20b
Release date
Aug 13, 2025
Last updated
Nov 25, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.04
Output token cost
$0.18

Limits

Output tokens
8,000 tokens
Context window
131,000 tokens

Latest news about openai/gpt-oss-20b

SiliconFlow

Official sourceBenchmark

Compare Hy3-preview and gpt-oss-20b across performance, cost, capabilities, and real-world use cases. See which model fits your needs.

SiliconFlow

Official sourceBenchmark

Compare Qwen3-VL-8B-Instruct and gpt-oss-20b across performance, cost, capabilities, and real-world use cases. See which model fits your needs.

Videos about openai/gpt-oss-20b

More models around openai/gpt-oss-20b