Model details
Wan v2.6 Text-to-Video
Wan v2.6 Text-to-Video is positioned as the production-grade entry in Alibaba Cloud's Wan video family, focused squarely on text-to-video generation rather than image-to-video or reference-driven workflows. Vercel AI Gateway exposes it under the model identifier alibaba/wan-v2.6-t2v, and developers reach it through the AI SDK's experimental generateVideo entry point by passing a plain text prompt such as a description of a mountain lake at sunrise. The same Gateway listing frames the model as a turnkey option for teams that want scripted, prompt-driven video synthesis without assembling their own pipeline.
The model's standout qualities are its end-to-end cinematic framing rather than raw clip length. Vercel documents automatic multi-shot scene composition, native audio baked into each generation, and resolutions up to 1080p, which together let a single prompt yield a short film-like segment instead of a flat shot. Output can reach roughly 15 seconds with continuous scenes and synchronized sound, a profile that suits storyboard prototyping, marketing teasers, product visualizations, and social clips where a coherent narrative arc matters more than minute-long runtime. A separate fal.ai page corroborates the 15-second HD output and native audio story, while the Vercel listing sets the lowest available configuration at $0.10 per second of generated video, giving production teams a straightforward per-second cost model to budget against.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- alibaba/wan-v2.6-t2v
- Release date
- Dec 16, 2025
- Last updated
- Dec 16, 2025
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 0 tokens
- Context window
- 0 tokens
Latest news about Wan v2.6 Text-to-Video
No articles yet. Fetch the latest news to show it here.