Google ships Gemini 3.5 Transcribe for real-time speech apps

The public-preview model handles 85+ languages, custom jargon and up to three speakers across live and recorded audio.

By · Published

Primary source: X

Why it matters

Google is packaging speech recognition as a native input layer for Gemini agents, reducing the plumbing needed to connect live audio, transcription and model-driven actions.

Google ships Gemini 3.5 Transcribe for real-time speech apps

Google released Gemini 3.5 Transcribe on August 26th, giving developers a dedicated speech-to-text model for live voice interfaces and recorded audio inside its Gemini developer stack.

https://x.com/sundarpichai/status/2092659467284517088

…

Reader comments

Conversation for this story loads after sign-in.