Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
SiliconFlow logo

Model details

ByteDance-Seed/Seed-OSS-36B-Instruct

The Seed-OSS-36B-Instruct comes from ByteDance's Seed Team, built as a 36-billion parameter language model that targets demanding workloads involving extended context windows, multi-step reasoning, and agentic task execution. The architecture supports a flexible thinking budget mechanism that lets developers control how the model allocates computation during responses. Setting the budget to -1 allows unlimited deliberation, while zero forces the model to output direct responses without intermediate reasoning. This design gives applications the ability to toggle between thorough analysis and quick answers depending on the task at hand.

The model was trained extensively on thinking intervals aligned to multiples of 512 tokens, which shapes how it structures reasoning chains during extended analysis. Benchmark comparisons show it outperforming alternatives in coding tasks, mathematical problem-solving, and overall intelligence metrics, while delivering faster response times. For deployment, the full 36B model in BF16 precision requires roughly 70GB of VRAM, but practitioners have successfully run quantized versions—EXL3, Q8, Q6_K, GPTQ, and AWQ formats—on consumer-grade hardware with dual GPUs, pushing context windows up to 300–350k tokens with stable performance. The Apache 2.0 license and OpenAI-compatible API make it accessible for teams building reasoning-heavy applications.

SiliconFlowByteDance-Seed/Seed-OSS-36B-Instructseed

Quick Info

Powered by
Provider
SiliconFlow
Model key
ByteDance-Seed/Seed-OSS-36B-Instruct
Release date
Sep 4, 2025
Last updated
Nov 25, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.21
Output token cost
$0.57

Limits

Output tokens
262,000 tokens
Context window
262,000 tokens

Transparent token rates

Compare ByteDance-Seed/Seed-OSS-36B-Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about ByteDance-Seed/Seed-OSS-36B-Instruct

SiliconFlow

CoverageBenchmark

BenchLeader's model page for ByteDance-Seed/Seed-OSS-36B-Instruct, dated 19 September 2026, provides a current third-party benchmark snapshot for the exact subject model. Reported scores include GPQA Diamond at 71.5% (Epoch AI) and 72.6% (Artificial Analysis), Humanity's Last Exam at 9.9%, Terminal-Bench Hard at 6.8%, The page also includes workload-level cost estimates based on list price, such as roughly $0.0003 for a 400/300 chat reply and about $0.015 for a 60,000/4,000 agentic coding session, with estimated completion times ranging from about 1.0 to 2.7 minutes. These figures position the model among current entries in BenchLea

Videos about ByteDance-Seed/Seed-OSS-36B-Instruct

More models around ByteDance-Seed/Seed-OSS-36B-Instruct