Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Impossibl logo

Model details

GPT OSS 20B

GPT OSS 20B is OpenAI's open-weight entry in the gpt-oss family, released as a mid-sized alternative to the larger 120B sibling and positioned for developers who want capable reasoning and agentic behavior without sending data to a closed API. Independent coverage frames the 20B and 120B pair as a new open-weight language model line aimed at versatile developer use cases, while the model's own description on the Ollama library emphasizes "powerful reasoning, agentic tasks, and versatile developer use cases" as its intended operating envelope.

On the technical side, the model carries 20.9 billion parameters and uses the gptoss architecture with MXFP4 quantization, which keeps the on-disk footprint to roughly 14 GB and makes it feasible to run locally on a single high-memory machine. The Ollama runtime exposes tools and a "thinking" mode around the weights, so practitioners can wire function-calling pipelines and chain-of-thought style reasoning directly into agents, with the package shipped under Apache License 2.0 terms. That combination of a compact parameter count, open weights, and tool-aware runtime makes GPT OSS 20B a practical fit for prototyping reasoning-heavy assistants, coding agents, and other task-oriented workflows that benefit from self-hosted inference.

Impossiblfireworks/gpt-oss-20bgpt-oss

Quick Info

Powered by
Provider
Impossibl
Model key
fireworks/gpt-oss-20b
Release date
Aug 5, 2025
Last updated
Aug 5, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.07
Output token cost
$0.30

Limits

Output tokens
32,768 tokens
Context window
131,072 tokens

Transparent token rates

Compare GPT OSS 20B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT OSS 20B

Impossibl

Coverage

The Baidu encyclopedia entry on gpt-oss-20b documents the model as an OpenAI open-weight release dated August 5, 2025, with 21 billion total parameters, 3.6 billion activated per token, and a MoE Transformer design that combines dense attention with a local banded sparse attention mechanism supporting a 128,000-token c The entry also notes Apache 2.0 licensing that permits parameter-level fine-tuning and commercial use, performance comparable to o3-mini on benchmarks including AIME and HealthBench, and downstream adoption milestones such as Microsoft making it available via Windows AI Foundry on August 6, 2025 and Silicon Flow launch

Videos about GPT OSS 20B

More models around GPT OSS 20B