Currently listed through these providers:
Model details
Seedance 2.0
Seedance 2.0 is ByteDance's generative video model designed for producing clips directly from text prompts or reference images, with third-party framing positioning it as a potential challenger to systems like Google's Veo 3.1, OpenAI's Sora 2, and Kuaishou's Kling 3.0. The official ByteDance product page describes its core as a unified multimodal audio-video joint generation architecture that accepts text, image, audio, and video as inputs, enabling rich multimodal content referencing and editing. This single-model architecture is meant to streamline workflows that previously required separate tools for conditioning, timing, and audio alignment, giving creators a more integrated path from prompt or image to finished clip.
ByteDance showcases Seedance 2.0's capabilities through SeedVideoBench-2.0, a multi-dimensional evaluation suite that scores the model on Text-to-Video, Image-to-Video, and Multimodal Task axes, with the company claiming a leading position across those dimensions. In practice, the model is best suited to creators who need flexible conditioning from text plus reference media, and to developers building editing or storytelling pipelines that benefit from combined text, image, audio, and video cues within one system. Because the benchmark leadership claim comes from ByteDance's own materials, the most practical takeaway is that Seedance 2.0 offers broad multimodal input breadth for video generation rather than independent confirmation of state-of-the-art performance.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- bytedance/seedance-2.0
- Release date
- Apr 14, 2026
- Last updated
- Apr 14, 2026
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 0 tokens
- Context window
- 0 tokens
Latest news about Seedance 2.0
No articles yet. Fetch the latest news to show it here.