News
StepAudio 2.5 TTS, a next-generation speech generation model, is launched by StepAudio.
StepAudio 2.5 TTS, a next-generation speech generation model, has been officially launched by StepAudio, featuring three core capabilities: global context control, text-based context control, and zero-shot reproduction. Users can precisely control details such as tone, rhythm, pauses, and emphasis in speech through natural language processing, achieving a leap from "reproducing voices" to "creating expressions." The model supports zero-shot reproduction of any timbre, generating high-quality speech without retraining.