What happened

Google has updated its Gemini Audio service with new Gemini 3.5 models, including Gemini 3.5 Live, 3.5 Live Experimental, and 3.5 Transcribe.

The new transcription capabilities automatically detect specialized jargon and handle more than 85 languages, according to the announcement.

Google says the models are built to deliver better precision for voice-controlled AI features, even when background noise or unclear speech is present.

Why it matters

Automatically stripping out filler words like ‘ums’ and ‘ahs’ can make AI-generated transcripts cleaner and more readable for meetings, notes, or captions.

Support for specialized jargon and a wide range of languages could make Gemini Audio more practical across fields like medicine, law, and international business.

Key facts

The update is part of Google’s Gemini Audio feature set.

It introduces three new models: Gemini 3.5 Live, 3.5 Live Experimental, and 3.5 Transcribe.

The transcription detects specialized jargon and supports more than 85 languages.

The models are designed to handle background noise and imperfect speech without struggling.

What to watch next

Whether the filler-word removal remains optional for users who need verbatim transcripts.

How well the jargon detection and multilingual support perform in real-world, noisy environments.

Sources