Google launches Gemini 3.8 Flash TTS
Google has unveiled Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, two new speech generation models designed for expressive character voicing and low-latency audio production. Accessible via Google AI Studio and the Gemini API, the models support promptable voice directing, multi-speaker dialogue, over 100 languages, and built-in SynthID watermarking.
Promptable natural-language voice directing represents the death knell for rigid SSML tags, positioning Google to mount serious pressure against standalone voice providers like ElevenLabs by integrating top-tier speech synthesis directly into its developer platform.
- –Natural-language prompting unlocks granular directorial control over character accents, pacing, and emotional shifts without requiring bespoke training runs or brittle markup.
- –The dual-model strategy sharply targets both ends of the voice stack: high-fidelity dramatic performance with Flash TTS and low-latency, high-volume throughput with Flash-Lite TTS.
- –Direct ecosystem integration across Google AI Studio, Gemini API, and Google Vids simplifies audio production pipelines for multimodal applications.
- –Built-in SynthID watermarking addresses growing enterprise and regulatory demands for provenance tracking in synthetic media.
DISCOVERED
1h ago
2026-09-24
PUBLISHED
7h ago
2026-09-24
RELEVANCE
AUTHOR
Sundar Pichai