Google Launches Gemini 3.8 Live Models With Native Speech-to-Speech and Background Tool Calls
Google announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking on September 15, 2026, marking a shift from traditional speech-to-text-to-speech pipelines toward native, continuous voice processing. The two models serve distinct purposes: the standard Live model prioritizes fluid, low-latency dialogue at scale, while the Extended Thinking variant handles complex multi-step reasoning and slow tool calls while narrating progress aloud. A key technical feature is asynchronous function calling, which allows background tasks to execute without interrupting the audio stream, reducing awkward silences in voice interactions. The models support over 97 languages and are being rolled out across Google surfaces including Gmail Live, Docs Live, and Search Live. Google has priced the Live API at $0.005 per minute for audio input and $0.018 per minute for output, making production-scale voice agent deployment more financially viable for product teams.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in