Gemini 3.8 text-to-speech says hello
The two text-to-speech models offer prompt-driven voice design across 100+ languages, 30-second voice replication with verbal consent checks, and line-by-line performance direction.
- Gemini 3.8 Flash TTS focuses on creative character direction, while Flash-Lite TTS is optimized for cost-efficient, high-volume dubbing and voice agents.
- Features include two-speaker script staging, long-form generation with minimal drift, and scripted non-verbal cues like laughs and gasps.
- Voice replication requires a matching verbal consent recording and embeds SynthID watermarks alongside C2PA credentials in generated audio.
- Available immediately in Google AI Studio and the Gemini API, alongside integrations in Gemini Notebook and Google Vids.
Developers and creators can now design custom synthetic voices from text prompts or clone voices with built-in consent verification.

Sources
Read this as text
Back to the AI news