Google launches Gemini 3.8 TTS models for custom, directed voices
Google has introduced Gemini 3.8 Flash TTS for detailed voice design and Flash-Lite TTS for high-volume audio generation across its AI, enterprise, notebook, and video products.

Key takeaways · 3
- 01
Use Flash TTS when a project requires original voices and detailed, line-by-line performance direction.
- 02
Use Flash-Lite TTS for high-volume dubbing, audio production, or expressive voice-agent workloads.
- 03
Teams can access the models through Google AI Studio, the Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.
Two models arrive
Google introduced Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS on September 23, 2026, calling them its most expressive audio-generation models yet. [1] Both models are available across Google AI Studio, the Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids, where users can generate custom character voices and direct scene dialogue. [1] Google says the releases turn voice generation from static presets into a dynamic creative studio. [1]
Creative control or scale
Gemini 3.8 Flash TTS creates voices from natural-language prompts and supports line-by-line direction over acting cues, pacing, dialect shifts, and backchanneling. [1] Gemini 3.8 Flash-Lite TTS is optimized for high-volume dubbing, audio content creation, and expressive voice agents, with fine-grained control over tone, pacing, and expressive nuance. [1] Google says the models include safety features intended to support responsible use. [1]
What it means
The product split makes Google’s intended trade-off explicit: Flash emphasizes original voice and performance design, while Flash-Lite emphasizes volume and cost efficiency. Within the Gemini Audio family, the TTS pair sits alongside 3.5 Live Translate, 3.5 Transcribe, 3.8 Live, and 3.8 Live Extended Thinking, extending the portfolio into generated performances. Availability through several Google products also provides multiple distribution paths. What the sources don't address: how pricing, latency, quality benchmarks, consent controls, or voice-cloning safeguards compare across the two models.
The release gives practitioners separate options for detailed voice creation and cost-efficient production at scale. Its availability across Google’s development and productivity products could simplify how teams incorporate generated speech into existing workflows.
Why it matters
Put this to work — one session a day, built for your industry.
Create a free account for a daily session — eight questions and one real-work challenge, on the news that affects your role.
Start freeHow this developed
23 September 2026
Google launches Gemini 3.8 TTS models for custom, directed voices
23 September 2026
Event created from source cluster.