The team still has work to do w/the current iteration of their model: it struggles w/some consonants, like "z," which led to inaccurate impressions of sounds like bees buzzing. They also can’t yet replicate how humans imitate speech, music, or sounds that are imitated
AI Model Struggles With Consonants in Speech Synthesis
By
–