Skip to main content
ModelsEnterpriseTools

Google’s new speech model Gemini 3.8 Live supports real-time reasoning

Source: SiliconANGLE (opens in a new tab) · Mike Wheatley

Intel Summary

Google LLC has launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, new voice processing models intended to reduce latency in voice-based artificial intelligence agents. The company bills them as its most advanced voice models to date, supporting near-real-time reasoning and simultaneous speech-and-thought processing.

Why It Matters

Latency remains a primary bottleneck for natural conversational agents. Integrating simultaneous reasoning into speech processing could enhance real-time responsiveness and performance for developer and enterprise voice agent applications.

Part of an ongoing development

Source

Google DeepMind releases Gemini 3.8 Live audio models

Google DeepMind released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two new speech-to-speech audio models for developers. The models reportedly lead the Artificial Analysis speech-to-speech leaderboard and are offered at $1.38 per hour of voice conversation to undercut OpenAI's GPT-Live-1. Claims are as reported; this summary makes no determination about accuracy or significance.

Organizations & Entities