Google launches Gemini 3.8 Live models optimized for real-time voice agents
Google has introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, positioning them as its most advanced live dialogue models to date.

Key takeaways · 3
- 01
Gemini 3.8 Live supports audio, image, video, and text inputs up to 128K tokens.
- 02
The models can output audio and text with a 64K token limit.
- 03
Google optimized the models for high-volume, latency-sensitive tasks like real-time dialogue.
New audio models
Google launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to enable voice agents more effectively. [1] The company describes them as its "most advanced live dialogue models yet." [1]
Collectively referred to as Gemini 3.8 Audio, these models are optimized for latency-sensitive, high-volume tasks like real-time dialogue. [2] Built on the Gemini 3 Pro architecture, the models accept audio, images, video, and text inputs with a context window of up to 128K tokens. [2] The models output audio and text with a token limit of up to 64K. [2]
What it means
The introduction of the Gemini 3.8 Live models underscores a shift toward native audio capabilities optimized for real-time responsiveness. [2] By offering a 128K token input window and outputting up to 64K tokens, the architecture supports extended, multimodal interactions. [2] What the sources don't address: How the latency and cost of these new models compare specifically to previous iterations in real-world deployment.
The release of Gemini 3.8 Live provides developers with models explicitly built for real-time, low-latency dialogue. The large multimodal input and output token windows allow for more complex and sustained interactions with voice agents.
Why it matters
Put this to work — one session a day, built for your industry.
Create a free account for a daily session — eight questions and one real-work challenge, on the news that affects your role.
Start freeHow this developed
16 September 2026
Event evidence refreshed from source cluster.
16 September 2026
Google launches Gemini 3.8 Live models optimized for real-time voice agents
16 September 2026
Event created from source cluster.
Sources
- Gemini 3.8 Audio (Live, Live Extended Thinking) - Model Card — Google DeepMinddeepmind.google
- Google launches Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its "most advanced live dialogue models yet", to more effectively enable voice agents (Google)Techmeme
- Google Launches New Gemini Models to Upgrade Enterprise Voice Agentsthe1news.com