Skip to content
Röst & Tal· LaunchAvailable

Google DeepMind unveils new models for realistic voice synthesis

Google DeepMind has unveiled the new text-to-speech models Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS for more realistic voice generation.

By the Aheadline editorial team·24 sep. 2026·2 min read·Source: Entity-watch: Google DeepMindVerifierad signalAI-generated
Google DeepMind unveils new models for realistic voice synthesis
Google DeepMind unveils new models for realistic voice synthesis
Google DeepMind unveils new models for realistic voice synthesis
By · Verktygs- & infrastrukturreporter
Last updated
Vad betyder det för mig?

What happened?

Google DeepMind has introduced two new text-to-speech models: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. The models allow for the creation of custom character voices and the management of dialects, emotional states, and scene instructions within direct dialogue. They are being integrated into tools such as Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.

Key facts

ModellerGemini 3.8 Flash TTS, Gemini 3.8 Flash-Lite TTS
PlattformarGoogle AI Studio, Gemini API, Google Vids

”Gemini 3.8 Flash TTS och Gemini 3.8 Flash-Lite TTS är våra mest uttrycksfulla ljudgenereringsmodeller hittills.”

— Google DeepMind, Utvecklare · Google Blog

Why it matters

The tools shift the boundary of voice synthesis from pre-set voices to fully customisable voice profiles. By enabling control over emotion and intonation, creators can generate more natural dialogues, while built-in safety features are intended to prevent misuse.

Who is affected?

The models are aimed at developers, audiobook creators, game developers, and podcast producers, as well as companies seeking to build more realistic voices into their applications.

Impact on the EU

The models are being released globally and are available via Google's platforms; however, use within the EU must comply with existing regulations regarding AI-generated content and data protection.

What else you should know

The launch marks a transition from static voice presets to more dynamic and controllable voice models for complex dialogue and storytelling.

Frequently asked questions

Quick answers about this story

Vad har hänt?
Google DeepMind har presenterat två nya text-till-tal-modeller: Gemini 3.8 Flash TTS och Gemini 3.8 Flash-Lite TTS.
När hände det?
Nyheten meddelades i samband med att Google offentliggjorde de nya röstmodellerna på sin officiella blogg.
Varför spelar det roll?
Modellerna gör det möjligt att skapa och finjustera realistiska AI-röster med känsla och dialekt, vilket förändrar hur ljudböcker, spel och poddar kan produceras.
Vilka plattformar får stöd?
Modellerna blir tillgängliga via Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook och Google Vids.
Original source
Entity-watch: Google DeepMind·blog.google

The link opens in a new window and leads to the publisher's own site.

Verifierad signal

Källan har spårats automatiskt från utgivaren via Aheadlines signalkedja.

AI-verktyg i artikeln

Topics

#Voice#Gemini#Models
[ STAY UP TO DATE ]

Get similar news straight to your inbox

No affiliate linksCancel anytimeGDPR-friendly
[ Frequency ]
[ What do you want to read about? ]

You'll receive updates on 2 topics.

The reader's room

Send in a question or an addition. The newsroom reads everything before it's published and replies when relevant. No AI-generated text – just people.

Sign in to submit a comment or question.

Loading comments…
How this affects you

Read the article through your role

  • Decide whether this affects strategy over 6–12 months or is just noise.
  • Discuss with leadership: do we own the right question or does ownership need to move?
  • Ask: what risk are we taking by NOT acting on this this quarter?

Generated angle — not editorial analysis of "Google DeepMind unveils new models for realistic voice synth"