Gemini 3.8 text-to-speech
Google is pitching Gemini 3.8 Flash TTS as a promptable voice studio, not just a preset speech model.
The release includes Flash TTS for custom voice design and Flash-Lite TTS for cheaper, higher-volume generation. Google says users can create voices from prompts, replicate an authorized voice from a 30-second sample, and direct dialogue line by line for pacing, emotion, accents, and conversational sounds. The models are rolling out in Gemini API and Google AI Studio, with consumer and enterprise placements split across Gemini Notebook, Google Vids, and Gemini Enterprise. Google also says generated audio is SynthID-watermarked, with consent verification for voice replication. HN · Frontpage AI's note
The release includes Flash TTS for custom voice design and Flash-Lite TTS for cheaper, higher-volume generation. Google says users can create voices from prompts, replicate an authorized voice from a 30-second sample, and direct dialogue line by line for pacing, emotion, accents, and conversational sounds. The models are rolling out in Gemini API and Google AI Studio, with consumer and enterprise placements split across Gemini Notebook, Google Vids, and Gemini Enterprise. Google also says generated audio is SynthID-watermarked, with consent verification for voice replication. HN · Frontpage AI's note
score 7