Google's new Flash TTS models let you design AI voices from scratch using text descriptions
Google has introduced Gemini 3.8 Flash TTS and Flash-Lite TTS, two text-to-speech models that support over 100 languages. These models allow users to generate custom voices using text descriptions, incorporate stage directions into spoken lines, and produce dialogue between two distinct voices from a single prompt.
Covered by 2 sources
- TThe Decoder↗Matthias Bastian18h ago
- MMarkTechPost↗Asif Razzaq16h ago