Google has introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, marking a major leap in real-time, audio-to-audio artificial intelligence. These models are built specifically to handle fluid, back-and-forth voice interactions while simultaneously processing complex underlying logic.
Core Capabilities
- Gemini 3.8 Live: Optimized for low-latency, natural dialogue. It allows users to speak organically with the AI, supporting real-time interruptions, dynamic tone adjustment, and fluid conversational flow without typical voice-assistant delays.
- Gemini 3.8 Live Extended Thinking: Tailored for complex tasks that require higher background reasoning. By leveraging parallel processing, the model executes multi-step logic and problem-solving “behind the scenes” while maintaining an uninterrupted, natural spoken conversation.
Key Benchmarks & Features
- Audio Intelligence: Achieves 97.7% on Big Bench Audio, establishing a new benchmark for spoken instruction comprehension and context retention.
- Continuous Audio Stream: Operates natively in an audio-to-audio framework rather than converting voice to text and back, drastically reducing response latency.
- Parallel Reasoning: Extended Thinking actively resolves edge cases, coding problems, or computational questions mid-conversation without pausing the verbal output.
These updates transform voice agents from simple command-execution tools into active collaborative partners. Developers can deploy Gemini 3.8 Live for live customer support, conversational tutoring, or real-time translation. Meanwhile, the Extended Thinking variant enables complex hands-free workflows—such as debugging code via voice or working through multi-variable strategy problems during a call—making voice interaction far more versatile across technical and professional fields.
We’re introducing Gemini 3.8 Live and 3.8 Live Extended Thinking – our best conversational AI.
The models talk, think, and handle tasks in the background without breaking your flow. 🧵 pic.twitter.com/nifcYx1C6L
— Google DeepMind (@GoogleDeepMind) September 15, 2026
You can read the full text of Google’s announcement at Blog.Google.
