Currently listed through these providers:
Model details
Step TTS 2
Step TTS 2 is referenced in third-party developer tooling as a legacy speech synthesis model from the StepFun ecosystem, with current community resources focusing on helping projects transition away from it toward the newer stepaudio-2.5-tts integration. The MCPMarket skill listing for StepFun TTS explicitly frames an automated migration path from Step-TTS-2 legacy models to the StepAudio 2.5 API, positioning the older model as a stepping stone rather than the active production target. This means developers evaluating the model today are more likely to encounter it through migration documentation than through fresh first-party capability descriptions.
Within the daymade/claude-code-skills repository, a stepfun-tts skill lives under the daymade-audio path and ships bundled scripts including tts_generate.py and an ab_compare.sh comparison helper, which suggests the model is being treated as a baseline for A/B testing against successors rather than a standalone production endpoint. The practical fit for Step TTS 2 today is therefore narrow: it serves as a reference point for developers porting older voice synthesis workflows forward, with multilingual voice generation, prosody handling, and batch processing now centered on the StepAudio 2.5 lineage in the documented toolchain.
Quick Info
Powered by- Provider
- StepFun (Global)
- Model key
- step-tts-2
- Release date
- Mar 1, 2026
- Last updated
- Jul 2, 2026
- Input modalities
- Output modalities
- Capabilities
- Base catalog fields only
Limits
- Output tokens
- 0 tokens
- Context window
- 0 tokens