Skip to main content
Aug 26

Google’s new AI transcription edits out your ‘ums’ and ‘ahs’

We got Gemini Audio models while we’re still waiting for the overdue Gemini 3.5 Pro launch. We got Gemini Audio models while we’re still waiting for

2 min read32 views1 tags
Originally reported bytheverge
We got Gemini Audio models while we’re still waiting for the overdue Gemini 3.5 Pro launch. We got Gemini Audio models while we’re still waiting for the overdue Gemini 3.5 Pro launch. Google has updated Gemini Audio with some new Gemini 3.5 models, introducing new transcription capabilities that automatically detect specialized jargon and more than 85 languages. Gemini 3.5 Live, 3.5 Live Experimental, and 3.5 Transcribe are designed to provide better precision for Google’s voice-controlled AI features, without struggling with background noise or when your speech is interrupted. Gemini 3.5 Transcribe is a completely new addition to the Gemini family, and its introduction comes as we’re still waiting for Google to release theGemini 3.5 Promodel that itpromised to roll out in June. Google says that 3.5 Transcribe “represents a major advancement from our previous transcription model, Chirp 3,” especially regarding multilingual performance and wording error rates. The transcription model allows users to “edit naturally with just your voice,” according to Google, and can automatically format text and remove filler words like “um” and “uh.” Users can provide a customized vocabulary to the model, allowing 3.5 Transcribe to automatically adapt transcription to unique spelling requirements and specialized jargon to prevent those words from being edited manually. It can also attribute speech for up to three speakers in pre-recorded audio, alongside providing word-level timestamps. Transcribe is launching alongside 3.5 Live and 3.5 Live Experimental, which build on the existing speech recognition tech that powersGemini’s voice chat mode. Gemini 3.5 Live is better at handling mid-sentence interruptions, language recognition, and live visual processing, while Gemini 3.5 Live Experimental goes further by narrating its progress step by step in real time while it tackles reasoning on more complex tasks. These Gemini Audio updates are rolling out starting today in English for all macOS Gemini app users, and the Rambler dictation feature on Android in select countries and languages. It’s also available for developers in public preview in the Gemini API via AI Studio and Antigravity. Google says that Chrome support is coming soon. A free daily digest of the news that matters most. This is the title for the native ad
#AI News
ES
Editorial StaffEditor

The Editorial Staff at AIChief is a team of professional content writers with extensive experience in AI and marketing. Founded in 2025, AIChief has quickly grown into the largest free AI resource hub in the industry.

View all posts
Reader feedback

What did you think of this story?

User Comments

Filter:
No comments yet. Be the first to comment!
Continue reading
View all news