TLDRocket
Sign in

Google releases Gemini 3.5 Transcribe, a new speech-to-text model for live and non-streaming transcription

Model release Confirmed 93% confidence first seen

Google has released Gemini 3.5 Transcribe, an AI speech-to-text model intended for real-time and application-based voice transcription. The coverage describes improvements such as reduced transcription errors and filler-word handling, with availability via Gemini API and AI Studio and rollout across Google products and voice features. Reported performance figures include lower word error rates for both streaming and non-streaming use cases, and the service is positioned as API-based rather than providing open weights.

Source coverage

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.