Google has introduced two new text-to-speech models to the #Gemini family: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. These models transform speech generation from rigid static presets into a highly dynamic creative studio. They enable creators, developers, and enterprises to build richer, more expressive audio experiences, while unlocking better user workflows in products like Gemini Notebook and Google Vids.
Gemini 3.8 Flash TTS is built specifically for deep creative direction and character design. It empowers users to generate entirely original voices from scratch using natural language prompts—bringing characters to life across gaming, immersive audiobooks, podcasts, and interactive media with fine-grained control over acting cues, pacing, and dialect shifts.
For high-volume, cost-efficient scale, Gemini 3.8 Flash-Lite TTS offers the perfect balance. Optimized for large-scale dubbing, high-throughput audio content creation, and highly responsive voice agents, it provides nuanced control over tone and expressiveness. These models complement Google's rapidly growing audio suite, joining 3.5 Live Translate, 3.5 Transcribe, and 3.8 Live.
This launch expands voice options from 30 default presets to an infinite generative library. Thanks to generative voice design in Gemini 3.8 Flash #TTS, creators can customize roles, accents, and vocal attributes across more than 100 languages and dialects, crafting anything from a dramatic, fantasy creature's growl to a charismatic narrator with a distinctive local cadence.
[AgentUpdate Depth Analysis] The introduction of the Gemini 3.8 TTS lineup marks a significant step forward in shifting voice user interfaces (VUI) from robotic speech synthesizers to fully expressive conversational AI Agents. By embedding natural language prompting into voice customization, Google lowers the engineering barrier to high-fidelity human-like vocal performances, challenging competitors like ElevenLabs and OpenAI. For the broader AI Agent ecosystem, realism and low-latency feedback are the ultimate frontiers. Gemini’s continuous progress in fine-grained pacing and acting cues will pave the way for virtual assistants, real-time tutors, and interactive NPCs that exhibit authentic emotional resonance, accelerating the next wave of agentic deployment across various industries.



