Google Launches Gemini 3.8 Live and 3.8 Live Extended Thinking for Real-Time Voice
Google launched two new real-time voice models: Gemini 3.8 Live, built for cost-efficient scale, and Gemini 3.8 Live Extended Thinking, aimed at complex, multi-step reasoning tasks conducted through live dialogue.
Details
- Gemini 3.8 Live: near real-time visual processing with contextual awareness, support for 97 languages with automatic mid-conversation switching, and background tool and API execution that doesn’t interrupt the dialogue
- Extended Thinking adds: simultaneous reasoning and speech for complex workflows, deeper analytical capabilities for enterprise tasks, and natural verbal acknowledgments like “Let me check that…” with live progress narration during multi-step operations
- Benchmarks: Extended Thinking scores 82.6 on Artificial Analysis’ Speech to Speech Quality Index, 68.6% on τ-Voice, and 97.7% on Big Bench Audio reasoning tasks; standard Live ranks second on Speech Agent Arena while leading on cost-effectiveness
- Availability: both launched September 15, 2026 via the Gemini API and AI Studio for developers, Gemini Enterprise in private preview for organizations, and consumer products including Search Live, Gemini Live, Workspace, Gmail, and Keep
What happened next
Splitting the Live lineup into a cost-optimized tier and a reasoning-focused Extended Thinking tier mirrors the text/flash-vs-reasoning split Google has used elsewhere in its model lineup, applied for the first time to real-time voice — giving developers a choice between cheap, fast voice interaction and slower, deeper reasoning delivered conversationally.