Skip to main content

Google launches Gemini 3.8 Live models optimized for real-time voice agents

16 SEPTEMBER 2026·2 MIN READ·3 SOURCES·Official source plus independent coverage

Google has introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, positioning them as its most advanced live dialogue models to date.

Google launches Gemini 3.8 Live models optimized for real-time voice agents

Key takeaways · 3

  • 01

    Gemini 3.8 Live supports audio, image, video, and text inputs up to 128K tokens.

  • 02

    The models can output audio and text with a 64K token limit.

  • 03

    Google optimized the models for high-volume, latency-sensitive tasks like real-time dialogue.

New audio models

Google launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to enable voice agents more effectively. [1] The company describes them as its "most advanced live dialogue models yet." [1]

Collectively referred to as Gemini 3.8 Audio, these models are optimized for latency-sensitive, high-volume tasks like real-time dialogue. [2] Built on the Gemini 3 Pro architecture, the models accept audio, images, video, and text inputs with a context window of up to 128K tokens. [2] The models output audio and text with a token limit of up to 64K. [2]

What it means

The introduction of the Gemini 3.8 Live models underscores a shift toward native audio capabilities optimized for real-time responsiveness. [2] By offering a 128K token input window and outputting up to 64K tokens, the architecture supports extended, multimodal interactions. [2] What the sources don't address: How the latency and cost of these new models compare specifically to previous iterations in real-world deployment.

The release of Gemini 3.8 Live provides developers with models explicitly built for real-time, low-latency dialogue. The large multimodal input and output token windows allow for more complex and sustained interactions with voice agents.

Why it matters
Daily session

Put this to work — one session a day, built for your industry.

Create a free account for a daily session — eight questions and one real-work challenge, on the news that affects your role.

Start free

How this developed

  1. 16 September 2026

    Event evidence refreshed from source cluster.

  2. 16 September 2026

    Google launches Gemini 3.8 Live models optimized for real-time voice agents

  3. 16 September 2026

    Event created from source cluster.

Sources

AI fluency, one session a day, built for your work.