Google’s new AI transcription edits out your ‘ums’ and ‘ahs’
Source: The Verge · Jess Weatherbed
Intel Summary
Google has introduced Gemini 3.5 Transcribe to its Gemini Audio suite, adding automatic speech recognition capabilities across more than 85 languages. The model is designed to recognize domain-specific jargon and automatically filter out speech disfluencies, such as filler words. The addition follows the release of Google's 3.5 Live Translate feature as the company continues expanding its Gemini 3.5 model portfolio ahead of anticipated larger foundation model releases.
Why It Matters
Automated filtering of filler words combined with multilingual jargon detection reduces post-processing overhead for meeting summaries, transcription workflows, and automated voice pipelines. The release enhances Google's native speech-to-text capabilities within the Gemini ecosystem, increasing competitive pressure on standalone transcription services and alternative speech models like Whisper.
Part of an ongoing development
SourceGoogle announces Gemini 3.5 Transcribe
Google has launched Gemini 3.5 Transcribe, an automated speech-to-text model supporting more than 85 languages. Google reports a 4.0 percent word error rate in streaming mode alongside a 70 percent reduction in latency compared to its previous Chirp 3 system. Claims are as reported; this summary makes no determination about accuracy or significance.
- Confidence
- High confidence
- Corroboration
- Corroborated
More coverage of this development
- Google's Gemini 3.5 Transcribe turns speech to text in 85 languages while auto-correcting your verbal stumblesThe DecoderIndependent reporting
- Google announces Gemini 3.5 Transcribe for AI-powered speech-to-textArs TechnicaIndependent reporting
Organizations & Entities
Related Intelligence
- ReportSame development
Google's Gemini 3.5 Transcribe turns speech to text in 85 languages while auto-correcting your verbal stumbles
Google has launched Gemini 3.5 Transcribe, an automated speech-to-text model supporting more than 85 languages. According to the company, the model performs real-time transcription with automatic filler word removal and speech stumble correction. Google reports a 4.0 percent word error rate in streaming mode alongside a 70 percent reduction in latency compared to its previous Chirp 3 system. The model also supports function calling to hand off transcription workflows directly to other Gemini models.
The Decoder - ReportSame development
Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text
Google has announced Gemini 3.5 Transcribe, a dedicated speech-to-text model designed to expand its voice transcription capabilities across consumer and enterprise products. According to reporting from Ars Technica, the underlying AI technology previously used in Gboard's Rambler feature will now be integrated into additional Google platforms, including the Chrome web browser. The release reflects Google's continued strategy of modularizing its Gemini model family for specific multimodal tasks like real-time audio transcription.
Ars Technica - DevelopmentNewAlso involving Google
Google DeepMind expands Co-Scientist to automate lab experiments and paper writing
Google DeepMind has expanded its Co-Scientist system from hypothesis generation into an integrated laboratory research platform. Powered by a Gemini-based multi-agent architecture, the system plans scientific experiments, interfaces directly with laboratory equipment to execute them, and writes corresponding scientific papers. Claims are as reported; this summary makes no determination about accuracy or significance.
1 reporting source - DevelopmentNewAlso involving Google
Google announced Gemini AI feature updates for Samsung devices
Google announced several Gemini AI integrations and feature updates for Samsung devices during Galaxy Unpacked 2026. The updates expand Google Gemini assistant functionality within the Android ecosystem, bringing context-aware features and automated task management directly onto partner consumer hardware. Claims are as reported; this summary makes no determination about accuracy or significance.