ModelsTools

Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text

Source: Ars Technica · Ryan Whitwam

Intel Summary

Google has announced Gemini 3.5 Transcribe, a dedicated speech-to-text model designed to expand its voice transcription capabilities across consumer and enterprise products. According to reporting from Ars Technica, the underlying AI technology previously used in Gboard's Rambler feature will now be integrated into additional Google platforms, including the Chrome web browser. The release reflects Google's continued strategy of modularizing its Gemini model family for specific multimodal tasks like real-time audio transcription.

Why It Matters

Expanding automated speech-to-text natively into widespread software like Chrome lowers the barrier for voice-first user interfaces and accessibility tools across the web. For developers and enterprise IT teams, dedicated audio models from hyperscalers increase competition against specialized transcription providers, potentially lowering transcription costs while embedding baseline automated transcription directly into everyday browser-based workflows.

Part of an ongoing development

Independent reporting

Google announces Gemini 3.5 Transcribe

Google has launched Gemini 3.5 Transcribe, an automated speech-to-text model supporting more than 85 languages. Google reports a 4.0 percent word error rate in streaming mode alongside a 70 percent reduction in latency compared to its previous Chirp 3 system. Claims are as reported; this summary makes no determination about accuracy or significance.

Confidence
High confidence
Corroboration
Corroborated

Organizations & Entities

Topics