Currently listed through these providers:
Model details
ByteDance-Seed/Seed-OSS-36B-Instruct
The model uses a 36 billion parameter architecture designed by ByteDance's Seed Team to handle extended reasoning and agentic workflows. It emphasizes long-context understanding alongside structured reasoning capabilities, featuring a distinctive thinking budget mechanism that lets developers tune inference-time compute. This approach supports variable reasoning depth depending on task complexity.
Post-training optimization through vLLM enables efficient deployment across inference scenarios. The thinking budget feature allows fine-grained control over reasoning length with values aligned to 512-token intervals, calibrated during training to balance performance against cost. Apache 2.0 licensing enables flexible integration, and OpenAI-compatible APIs facilitate straightforward adoption for developers building agents, automation systems, or complex problem-solving applications.
Quick Info
Powered by- Provider
- Nvidia
- Model key
- bytedance/seed-oss-36b-instruct
- Release date
- Sep 4, 2025
- Last updated
- Nov 25, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 262,000 tokens
- Context window
- 262,000 tokens
Latest news about ByteDance-Seed/Seed-OSS-36B-Instruct
Videos about ByteDance-Seed/Seed-OSS-36B-Instruct
Recent tweets and retweets from Nvidia
More models around ByteDance-Seed/Seed-OSS-36B-Instruct
This exact model name is also listed by 2 other providers.