Google DeepMind Launches Two Gemini 3.8 Live Audio Models for Real-Time Voice AI
Google DeepMind has unveiled two new audio-focused AI models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, designed to enable more natural, real-time voice conversations with AI systems. The standard model targets high-volume, cost-effective voice interactions with near real-time responsiveness, while the Extended Thinking variant handles complex reasoning and can coordinate multiple background agents during an active conversation. Both models go beyond existing Gemini Audio capabilities such as transcription, translation, and text-to-speech by focusing on the conversational layer itself. They are accessible through Google AI Studio, the Gemini API, Gemini Live API, and related Google services, though no public pricing or performance benchmarks have been published. Google also notes that SynthID watermarking is applied across its audio work to identify AI-generated or edited speech.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in