Radio
Now Playing
Quickyla Radio โ€” Click to play
Open โ†’
3 min left
Back to News

Google rolls out Gemini Audio to improve real-time dialogue and speech recognition

Itโ€™s been a few months since Google debuted the first of its Gemini 3.5 models . With the rapid pace of AI development, the company has since released Gemini 3.6 and Gemini 3.7 . However, the tech giโ€ฆ

Google rolls out Gemini Audio to improve real-time dialogue and speech recognition
Android Authority โ€” 26 August 2026
Text:
1 0 0

Affiliate links on Android Authority may earn us a commission. Learn more.

Itโ€™s been a few months since Google debuted the first of its Gemini 3.5 models . With the rapid pace of AI development, the company has since released Gemini 3.6 and Gemini 3.7 . However, the tech giant hasnโ€™t completely moved on from the 3.5 family just yet, as it is introducing new members to the group today.

Google has announced a trio of new speech-related models: Gemini 3.5 Transcribe, Gemini 3.5 Live, and Gemini 3.5 Live Experimental. Together, these models make up what the company calls Gemini Audio. According to Google, these models were built to power real-time dialogue and speech recognition for more natural and responsive conversations.

Gemini 3.5 Transcribe is described as the Mountain View-based firmโ€™s most precise speech-to-text model yet. It will replace the previous transcription model, Chirp 3. Google claims that Gemini 3.5 Transcribe offers better precision, deep context awareness, and automatic language detection across more than 85 languages.

Weโ€™ve already had a taste of this model through the new Rambler feature for Gboard, which is available in select countries and languages. Users will also find Gemini 3.5 Transcribe on Google Antigravity, and the Gemini app on macOS. Eventually, it will also head to Chrome, allowing you to talk to type in any web field.

As for Gemini 3.5 Live, this model is designed to handle mid-sentence interruptions and process live visuals. Additionally, it can blend multiple languages and trigger background tools. The model is capable of doing all of this without the need to pause the conversation.

Meanwhile, Live Experimental takes things a little further. This model is meant to handle more complex tasks, reasoning directly while speaking and narrating its progress step by step.

Youโ€™ll be able to start trying out the Gemini Audio family of models soon. Google says that these models are rolling out for everyone in Search Live, Gemini Live, Docs, Keep, Gmail, the Gemini app, and Gboard. Meanwhile, theyโ€™ll be available across the Gemini API via Google AI Studio and Google Antigravity for developers. And for enterprise customers, these models will come to the Gemini Enterprise Agent Platform and Gemini Enterprise for Customer Experience.

Read Full Story at Android Authority โ†’
Advertisement
React:
Sponsored

More to Read

Google announces Tensor G6, bringing 4K Portrait Video and โ€ฆ
๐Ÿ’ป Technology
Google announces Tensor G6, bringing 4K Portrait Video and more to the Pixel 11 series
Android Authority ยท 14 days ago
Apple announces Ultra 4 and Series 12 Watch models for nextโ€ฆ
๐Ÿ’ป Technology
Apple announces Ultra 4 and Series 12 Watch models for next month
9to5Mac ยท 12 days ago
Sony says the 3.5mm headphone jack is still a hit, and its โ€ฆ
๐Ÿ’ป Technology
Sony says the 3.5mm headphone jack is still a hit, and its phone sales back it up
Android Authority ยท 13 days ago
Iran voids 60-day nuclear negotiation deadline with US
๐ŸŒ World News
Iran voids 60-day nuclear negotiation deadline with US
France 24 ยท 8 days ago
Idlib residents celebrate court's death sentence for Assad,โ€ฆ
โš”๏ธ War & Conflict
Idlib residents celebrate court's death sentence for Assad, former officials
Al Jazeera ยท 14 days ago
Iran demands U.S. recognize defeat in Strait of Hormuz tensโ€ฆ
๐Ÿ›๏ธ Politics
Iran demands U.S. recognize defeat in Strait of Hormuz tensions
NBC News ยท 10 days ago
Full view