Currently listed through these providers:
Model details
Seed 1.6 Vision
Seed 1.6 Vision belongs to ByteDance's Seed family of multimodal models, designed to process combined image and text inputs while producing text-based outputs. Independent gateway documentation indicates that the Seed 1.6 line distinguishes itself by offering configurable reasoning behavior, letting callers toggle between thinking and non-thinking modes depending on whether deeper deliberation or faster response is desired for a given task. This dual-mode design reflects a broader trend in frontier multimodal systems toward giving developers explicit control over the inference-time reasoning budget.
Positioned within ByteDance's vision-capable model family, Seed 1.6 Vision is suited for practical workloads that require grounded interpretation of visual content alongside natural language understanding, such as document analysis, visual question answering, and structured extraction from images. Third-party routing listings expose the broader Seed 1.6 family with a large that quick-info value window and a capability profile spanning vision input, tool use, reasoning, and structured output, making the vision variant a reasonable choice for agentic pipelines and multimodal assistants that need to chain perception with downstream tool calls or JSON-constrained responses.
Quick Info
Powered by- Provider
- Volcengine Ark
- Model key
- doubao-seed-1-6-vision-250815
- Release date
- Aug 15, 2025
- Last updated
- Aug 15, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.11875
- Output token cost
- $1.18747
Limits
- Output tokens
- 32,000 tokens
- Context window
- 256,000 tokens