Model details
HappyHorse 1.1 Text-to-Video
HappyHorse 1.1 Text-to-Video is positioned as a cinematic generative video model that turns written prompts into short, motion-rich clips, with reseller pages describing its intended use as creative generation featuring dynamic details and stronger subject consistency. It ships alongside two sibling variants for image-to-video and reference-to-video, all reached through a shared endpoint where only the model parameter changes, which makes it straightforward to mix modalities inside the same integration. On reseller playgrounds, creators can steer the output with configurable duration from three to fifteen seconds, resolutions spanning 480p, 720p, and 1080p, and a broad set of aspect ratios, giving practical control over the final look without leaving the text-to-video flow.
Beyond the creative framing, the model is aimed at workloads that benefit from realistic motion and smoother subject behavior, including short-form narrative clips, product showcases, and avatar-style content where audio-visual synchronization matters. Per-second billing on both EvoLink and Crun lets users scale cost with clip length, and predictable quality tiers at 720p and 1080p make it easy to pick a fidelity target for a given project. Because the same endpoint covers all three HappyHorse 1.1 variants, teams can prototype with text-to-video and then graduate to image- or reference-driven generation as their pipeline matures, keeping prompt design and quality settings consistent across modes.
Quick Info
Powered by- Provider
- Alibaba Token Plan
- Model key
- happyhorse-1.1-t2v
- Release date
- Jul 17, 2026
- Last updated
- Jul 17, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 0 tokens
- Context window
- 0 tokens
Latest news about HappyHorse 1.1 Text-to-Video
No articles yet. Fetch the latest news to show it here.