Google DeepMind introduced two new AI models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, on September 17, 2026. These models are designed to advance near real-time reasoning for voice agents, making AI conversations more intuitive and intelligent.
Gemini 3.8 Live is built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking is designed for high-complexity tasks, offering increased intelligence and multi-step reasoning. It achieved the top spot on Artificial Analysis' Speech to Speech Quality Index with 82.6 and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark.
These models provide developers and enterprises with building blocks for reliable, production-ready voice agents. They also aim to make speaking with Gemini more fluid and collaborative across the Gemini app, Google Workspace, and Search, enabling users to tackle complex tasks using voice.
The Gemini 3.8 Live models approach real-time voice agent functionality differently from other offerings. Gemini 3.8 Live Extended Thinking keeps reasoning within the voice model, allowing it to continue speaking while executing asynchronous tool calls. This contrasts with approaches that separate real-time conversation from backend reasoning, pushing more orchestration to the application layer.
✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →
One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.
One email a day. Unsubscribe in one click, any time.
Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.
▶ Play today's briefNew every morning, and the back catalogue is archived by date.
Google's Gemini has released two new text-to-speech (TTS) models, 3.8 Flash TTS and 3.8 Flash-Lite TTS, expanding its audio family. These models offer advanced voice generation capabilities for creative applications and high-volume audio content, providing more dynamic and expressive audio experiences for developers and enterprises.
Google released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, new voice models designed to reduce latency in voice agents by allowing them to continue speaking during background processing. This release follows OpenAI's GPT-Live-1, with both companies offering different architectural approaches to address the latency challenge in conversational AI.
Google has introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two new AI models designed to improve near real-time reasoning for voice agents. These models aim to make AI conversations more intuitive and intelligent, offering capabilities for both scalable, cost-efficient applications and high-complexity tasks requiring multi-step reasoning.