Google ships Gemini 3.5 Transcribe for real-time speech apps
The public-preview model handles 85+ languages, custom jargon and up to three speakers across live and recorded audio.
By Ryan Merket · Published
Primary source: X
Why it matters
Google is packaging speech recognition as a native input layer for Gemini agents, reducing the plumbing needed to connect live audio, transcription and model-driven actions.

Google released Gemini 3.5 Transcribe on August 26th, giving developers a dedicated speech-to-text model for live voice interfaces and recorded audio inside its Gemini developer stack.
https://x.com/sundarpichai/status/2092659467284517088
…