MIT CSAIL, JHU, & USC researchers have developed "CapSpeech," a text-to-speech framework that generates voices w/controllable timbre & speaking style via text prompts.
— MIT CSAIL (@MIT_CSAIL) 5 juin 2025
Imagine designing a voice just by describing it. CapSpeech can customize age, timbre, accent, emotion, & more:… pic.twitter.com/Dw0i6JUG19
MIT CSAIL, JHU, & USC researchers have developed "CapSpeech," a text-to-speech framework that generates voices w/controllable timbre & speaking style via text prompts. Imagine designing a voice just by describing it. CapSpeech can customize age, timbre, accent, emotion, & more: