Google Launches Gemini 3.5 Transcribe with Voice-Driven Workflows for macOS
Google announced Gemini 3.5 Transcribe on August 26, 2026, describing it as its most precise speech-to-text model to date. The model goes beyond basic dictation by combining voice commands with on-screen context to summarize local files, repurpose text across apps, and generate images at the cursor. It can call other Gemini models in the background, allowing a single spoken instruction to trigger multi-step AI workflows without manual app-switching. The model also supports multilingual transcription in over 85 languages, background-noise robustness, and speaker attribution for pre-recorded audio. In addition to the macOS Gemini app, Gemini 3.5 Transcribe is available to developers via API.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in