We're sharing our learnings from a small-scale preview of Voice Engine, a model which uses text input and a single 15-second audio sample to generate natural-sounding speech that closely resembles the original speaker.
OpenAI Voice Engine: Text-to-Speech AI Model Preview
By
–