Gemini 3.8 text-to-speech says hello
Today, we’re introducing two new text-to-speech models to the Gemini family, transforming voice generation from static presets into a dynamic creative studio.
Key points
- Gemini 3.8 Flash TTS: Built for deep creative direction and character design.
- Create entirely new voices from scratch using natural language prompts to bring characters to life across gaming, immersive audiobooks, podcasts, and interactive media.
- These models complement our fast-growing Gemini Audio family, following 3.5 Live Translate, 3.5 Transcribe, 3.8 Live, and 3.8 Live Extended Thinking.
- Generative voice design: With Gemini 3.8 Flash TTS, create bespoke voices from scratch by customizing role, accent and voice characteristics across more than 100 languages and dialects using natural language prompting — whether you're bringing a dramatic, fire-breathing dragon to life or crafting a charismatic narrator with a distinct regional cadence.
Sources (1)
- [1]Gemini 3.8 text-to-speech says helloGoogle DeepMind Blog · Sep 23, 03:25 PM
Today, we’re introducing two new text-to-speech models to the Gemini family, transforming voice generation from static presets into a dynamic creative studio.
- Gemini 3.8 Flash TTS: Built for deep creative direction and character design.
Extractive summary: sentences quoted from the sources.