Intelligent transcription with Gemini 3.5 Transcribe
Google ● Covered by 6 sources
Google introduced Gemini 3.5 Transcribe, a speech-to-text model aimed at real-time, intelligent transcription and voice interactions in apps and developer tools. It reports a 4.0% average Word Error Rate (WER) for streaming and 2.6% for non-streaming use-cases, based on Artificial Analysis. Availability expands through Gemini API (Live API and Interactions API) and Google AI Studio, plus public preview support across products like Gboard/Rambler on Android, the Gemini macOS app, and upcoming Chrome voice dictation.
Why it matters
Now you can get more intelligent speech-to-text transcription with Gemini 3.5 Transcribe.