Gemini 3.8 text-to-speech says hello
Google has released Gemini 3.8, introducing native text-to-speech capabilities that allow the model to generate spoken audio directly alongside text responses. By bypassing the need for separate voice synthesis software, this integration aims to improve latency and emotional nuance in conversational AI interactions. This development follows a broader industry trend toward multimodal models that process and output multiple media formats simultaneously. Users can now access these features through Google's developer platforms to build more responsive and natural-sounding AI applications.
Covered by 2 sources
- GGoogle DeepMind Blog↗21h ago
- HHacker News↗swolpers21h ago