Google DeepMind unveils new models for realistic voice synthesis
Google DeepMind has unveiled the new text-to-speech models Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS for more realistic voice generation.

What happened?
Google DeepMind has introduced two new text-to-speech models: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. The models allow for the creation of custom character voices and the management of dialects, emotional states, and scene instructions within direct dialogue. They are being integrated into tools such as Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.
Key facts
| Modeller | Gemini 3.8 Flash TTS, Gemini 3.8 Flash-Lite TTS |
|---|---|
| Plattformar | Google AI Studio, Gemini API, Google Vids |
”Gemini 3.8 Flash TTS och Gemini 3.8 Flash-Lite TTS är våra mest uttrycksfulla ljudgenereringsmodeller hittills.”
Why it matters
The tools shift the boundary of voice synthesis from pre-set voices to fully customisable voice profiles. By enabling control over emotion and intonation, creators can generate more natural dialogues, while built-in safety features are intended to prevent misuse.
Who is affected?
The models are aimed at developers, audiobook creators, game developers, and podcast producers, as well as companies seeking to build more realistic voices into their applications.
Impact on the EU
The models are being released globally and are available via Google's platforms; however, use within the EU must comply with existing regulations regarding AI-generated content and data protection.
What else you should know
The launch marks a transition from static voice presets to more dynamic and controllable voice models for complex dialogue and storytelling.
Quick answers about this story
Vad har hänt?
När hände det?
Varför spelar det roll?
Vilka plattformar får stöd?
The link opens in a new window and leads to the publisher's own site.
Källan har spårats automatiskt från utgivaren via Aheadlines signalkedja.
AI-verktyg i artikeln
Topics
Get similar news straight to your inbox
The reader's room
Send in a question or an addition. The newsroom reads everything before it's published and replies when relevant. No AI-generated text – just people.
Sign in to submit a comment or question.
Read the article through your role
- Decide whether this affects strategy over 6–12 months or is just noise.
- Discuss with leadership: do we own the right question or does ownership need to move?
- Ask: what risk are we taking by NOT acting on this this quarter?
Generated angle — not editorial analysis of "Google DeepMind unveils new models for realistic voice synth"