I ran the same prompt for a London Estuary accent, a Newcastle accent and an Exeter, Devon accent – all three audio files are now embedded in my blog post
MULTIMODAL AI
-

Google Gemini Flash TTS Model Demonstrates Advanced Text-to-Speech
By
–
The example prompt for Google's new Gemini Flash TTS text-to-speed model is a lot https://
simonwillison.net/2026/Apr/15/ge
mini-31-flash-tts/
… -

Physion-Eval: Benchmarking Physical Realism in AI-Generated Videos
By
–
How physically realistic are our AI-generated videos? Physion Labs, Stanford University, MIT, Harvard University, and Character AI introduce Physion-Eval. This new benchmark uses expert human reasoning to meticulously diagnose and explain physical realism failures in
-

OCR Models Struggle with Non-Latin Unicode Scripts Benchmark
By
–
GlotOCR Bench OCR Models Still Struggle Beyond a Handful of Unicode Scripts paper: https://
huggingface.co/papers/2604.12
978
… -
Single Image Transforms Into Immersive AI World
By
–
Turns a single image into a world. https://t.co/LVVy3lgmgq
— Robert Scoble (@Scobleizer) 15 avril 2026Turns a single image into a world.
-

Gemini 3.1 Flash TTS: Advanced Speech Synthesis Across 70 Languages
By
–
Gemini 3.1 Flash TTS released.
— Chubby♨️ (@kimmonismus) 15 avril 2026
Highlight: highly controllable speech via simple text prompts, more natural voices, and support for 70+ languages. Nice one https://t.co/4V8MhXmIxc pic.twitter.com/MsvCMdWhLRGemini 3.1 Flash TTS released. Highlight: highly controllable speech via simple text prompts, more natural voices, and support for 70+ languages. Nice one
-

Gemini 3.1 Flash TTS Audio Tags Flexibility Explored
By
–
[excitedly] I've been having fun with Gemini 3.1 Flash TTS, the audio tags are really flexible, you can do so much with them.
— fofr (@fofrAI) 15 avril 2026
[like dracula] I can't believe things like that just work. https://t.co/847vOEP1kd pic.twitter.com/PpfKMB9Dsw[excitedly] I've been having fun with Gemini 3.1 Flash TTS, the audio tags are really flexible, you can do so much with them. [like dracula] I can't believe things like that just work.
-

Google Gemini 3.1 Flash Shows Strong Progress on TTS
By
–
The progress from 2.5 to 3.1 has been super strong! Excited to see this much improvement for a Flash model. Learn more in our blog: https://
blog.google/innovation-and
-ai/models-and-research/gemini-models/gemini-3-1-flash-tts/
… -
Gemini 3.1 Flash TTS transforms scripts into studio-quality narration
By
–
Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio
. Whether you’re creating a pitch deck or recording a passion project, transform your scripts into studio-quality narration: -

Gemini 3.1 Flash TTS: Advanced Text-to-Speech Model Launches
By
–
Introducing Gemini 3.1 Flash TTS 🗣️, our latest text to speech model with scene direction, speaker level specificity, audio tags, more natural + expressive voices, and support for 70 different languages.
— Logan Kilpatrick (@OfficialLoganK) 15 avril 2026
Available via our new audio playground in AI Studio and in the Gemini API! pic.twitter.com/5PpBdhQMNgIntroducing Gemini 3.1 Flash TTS , our latest text to speech model with scene direction, speaker level specificity, audio tags, more natural + expressive voices, and support for 70 different languages. Available via our new audio playground in AI Studio and in the Gemini API!