Back
AI summary
Written by AI from the official notes. Check them for exact details.Google has released Gemini 3.8 Flash TTS and Flash-Lite TTS models for text-to-speech applications.
- Gemini 3.8 Flash TTS offers studio-grade voice fidelity and nuanced acting.
- Flash-Lite TTS is designed for high-throughput production and real-time applications.
- Custom vocal personas can be created from text prompts.
- Voice replication with consent verification is now available.
- Access to 150+ prebuilt and custom voices is included.
Why it matters: Developers and businesses using TTS technology should explore these new models for enhanced voice applications.
Full release notes3 changes
- Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS generally available (GA): Released our next-generation text-to-speech (TTS) audio models and the Gemini API Voices endpoint (
/v1beta/voices): Gemini 3.8 Flash TTS (gemini-3.8-flash-tts): Flagship creative TTS model engineered for studio-grade voice fidelity, nuanced acting, regional dialects, and long-form multi-turn stability. - Gemini 3.8 Flash-Lite TTS (
gemini-3.8-flash-lite-tts): Fast, cost-efficient TTS model built to replacegemini-3.1-flash-tts-previewfor high-throughput production and real-time voice agent cascades. - Voice design, Voice replication, and the Extended Voice Library: Create persistent custom vocal personas from text prompts, replicate voices with consent verification, and query 150+ prebuilt and custom voices.
See the Text-to-speech guide to get started.