OpenAI makes you call sales for a custom voice. Google just made it self-serve.
The New Stack Amanda Caswell ● Covered by 9 sources
Google released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, adding prompt-based voice design plus voice replication via an API Voices endpoint. Google requires two recordings from the same speaker—10 to 30 seconds each—plus a separate consent recording, and returns a voice_id kept in the project for one year. Developers can now create and reuse custom synthetic voices self-serve through the Gemini API and AI Studio instead of relying on a fixed set of voices, while existing prompts with stage directions may break due to changed transcript handling.
Why it matters
Google released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS today through the Gemini API and Google AI Studio. The post OpenAI makes you call sales for a custom voice. Google just made it self-serve. appeared first on The New Stack.
Related stories
Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents
MarkTechPost · 1 week ago ·
8
Improved Gemini audio models for powerful voice experiences
Google DeepMind · 9 months ago ·
31