Model details
Stable Audio 2.5 (Text-to-Audio)
Stable Audio 2.5 is the latest audio model from Stability AI, distributed through the fal platform and positioned as the first in the line designed specifically with enterprise-grade creative workflows in mind. Rather than serving purely experimental users, it targets brand, advertising, and production teams that need commercially ready audio with room for creative control. The release emphasizes a leap forward in both output quality and steerability, aiming to close the gap between custom sound identity demand and the small share of creative work that currently uses branded audio at scale.
The model family offers multiple interaction modes for different creative entry points, including text-driven generation as well as audio-to-audio transformation and audio inpainting, letting producers refine existing material rather than start from silence. Within the broader Stable Diffusion model family index, these audio capabilities sit alongside long-standing text-to-image and inpainting tools, giving teams a familiar workflow context when moving between visual and sonic tasks. Practical fit centers on short, brand-aligned music and sound-design snippets where consistent identity and repeatable quality matter more than long-form composition.
Quick Info
Powered by- Provider
- DigitalOcean
- Model key
- fal-ai/stable-audio-25/text-to-audio
- Release date
- Oct 8, 2025
- Last updated
- Apr 16, 2026
- Input modalities
- Output modalities
- Capabilities
- Base catalog fields only
Limits
- Output tokens
- 0 tokens
- Context window
- 0 tokens
Latest news about Stable Audio 2.5 (Text-to-Audio)
No articles yet. Fetch the latest news to show it here.