Skip to content
Röst & Tal· LaunchAvailable

Google unveils Gemini 3.5 Transcribe for smarter speech-to-text

Google is launching Gemini 3.5 Transcribe, a new speech-to-text AI model that removes filler words and supports more than 85 languages.

By the Aheadline editorial team·30 aug. 2026·2 min read·Source: Google News: Gemini feature (en)Aggregerad källaAI-generated
Google unveils Gemini 3.5 Transcribe for smarter speech-to-text
Google unveils Gemini 3.5 Transcribe for smarter speech-to-text
By · Verktygs- & infrastrukturreporter
Last updated

What happened?

Google has announced Gemini 3.5 Transcribe, a new AI model focused on audio transcription and speech-to-text. The model supports automatic detection of more than 85 languages, as well as specialised jargon and industry terminology. Beyond basic text conversion, the model has built-in functionality to automatically edit out hesitation words such as "um" and "ah."

Key facts

ModellnamnGemini 3.5 Transcribe
SpråkstödÖver 85 språk
LanseringsmånadJuli 2026

Why it matters

Standard speech-to-text often requires extensive post-processing to remove filler words and correctly identify technical concepts. By integrating these features directly into Gemini 3.5 Transcribe, developers building voice-based AI agents can reduce both complexity and latency.

Who is affected?

Developers, companies, and organisations building applications for audio analysis, meeting notes, and automated transcription are directly impacted. End-users of services integrating Google Cloud Vertex AI will also gain access to more precise and refined text from audio files.

Impact on the EU

The model is distributed globally via Google Cloud Vertex AI and AI Studio, making it available to developers in the EU. However, companies processing audio data within the EU must ensure compliance with GDPR and the EU AI Act regarding data storage and privacy.

What else you should know

The announcement coincides with updates to several other models in the Gemini family, as Google continues to specialise its AI models for specific infrastructure needs.

Frequently asked questions

Quick answers about this story

Vad har hänt?
Google presenterade den nya AI-modellen Gemini 3.5 Transcribe för automatisk tal-till-text och ljudtranskribering.
När hände det?
Offentliggörandet skedde i juli 2026.
Varför spelar det roll?
Modellen förenklar utveckling av röstbaserade tjänster genom att automatiskt rensa fyllnadsord och hantera över 85 språk samt facktermer direkt i modellen.
Hur blir modellen tillgänglig?
Modellen görs tillgänglig för utvecklare via Googles utvecklarplattformar Google AI Studio och Vertex AI.
Original source
Google News: Gemini feature (en)·news.google.com

The link opens in a new window and leads to the publisher's own site.

Aggregerad källa

Källan är en aggregator eller syndikering — vi rekommenderar att verifiera hos primärutgivaren.

AI-verktyg i artikeln

Topics

#AI-verktyg#Voice#Large Language Models (LLMs)#Gemini 3.5 Flash
[ STAY UP TO DATE ]

Get similar news straight to your inbox

No affiliate linksCancel anytimeGDPR-friendly
[ Frequency ]
[ What do you want to read about? ]

You'll receive updates on 2 topics.

The reader's room

Send in a question or an addition. The newsroom reads everything before it's published and replies when relevant. No AI-generated text – just people.

Sign in to submit a comment or question.

Loading comments…
How this affects you

Read the article through your role

  • Decide whether this affects strategy over 6–12 months or is just noise.
  • Discuss with leadership: do we own the right question or does ownership need to move?
  • Ask: what risk are we taking by NOT acting on this this quarter?

Generated angle — not editorial analysis of "Google unveils Gemini 3.5 Transcribe for smarter speech-to-t"