3 Sources
[1]
The Gemini 3.8 family is getting a little bigger with two new additions
It's been about two weeks since Google announced Gemini 3.8 Flash, which arrived only three weeks after it launched Gemini 3.7 Flash. The time for a new successor hasn't come quite yet, but we are getting introduced to new AI models today. While we wait for the next step up, Google is expanding the
[2]
Gemini 3.8 Live Extended Thinking powers Gemini Live, Gmail, & Keep
Google today announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking as its "most advanced live dialogue models yet." Following the launch of 3.1 Flash Live in March, the new models aim to "make conversing with AI feel more intuitive and intelligent." As the name suggests, Gemini 3.8 Live
[3]
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
Today, we're introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent. * Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with
Share
Copy Link
Google unveiled two new AI models expanding the Gemini 3.8 family: Live and Live Extended Thinking. These production-ready voice assistants bring near real-time reasoning and conversational intelligence to Google Workspace, Search, and the Gemini app, with Live Extended Thinking capturing the top spot on Artificial Analysis' Speech to Speech Quality Index at 82.6.
Google is expanding its Gemini 3.8 family with two new AI models designed to make conversational AI more intuitive and capable. The company announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, positioning them as production-ready voice assistants that enhance how users interact with AI across Google's ecosystem.
1
2
The two models serve distinct purposes within Google's AI strategy. Gemini 3.8 Live is built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and real-time visual input processing. In contrast, Gemini 3.8 Live Extended Thinking targets high-complexity tasks requiring multi-step reasoning and increased intelligence.
3
Both live dialogue models represent Google's push toward making voice agents more reliable and natural in their interactions.Gemini 3.8 Live processes visual inputs in near real-time while handling tools and API calls in the background. This allows the model to acknowledge requests and maintain conversation flow simultaneously. The model can detect and transition between 97 supported languages mid-conversation, making it adaptable for global users.
1
This capability positions it as a practical solution for developers and enterprises seeking cost-effective voice agents that can scale across diverse user bases.Gemini 3.8 Live Extended Thinking delivers what Google calls simultaneous reasoning and speaking, a feature that enables the model to think through complex problems while maintaining conversational flow. The model captured the top spot on Artificial Analysis' Speech to Speech Quality Index with a score of 82.6.
2
3
It also leads in agentic task completion with 68.6% on τ-Voice and achieved 35.1% on Sierra's τ-Voice-banking benchmark, which tests real-time, audio-native conversational AI agents on complex banking and fintech customer support tasks.1
The model scored 97.7% on Big Bench Audio, demonstrating strong reasoning capabilities while maintaining competitive pricing compared to other frontier models.
2
Meanwhile, Gemini 3.8 Live secured second place in the Speech Agent Arena, showing high user preference alongside its cost-effectiveness.3
Both models are rolling out across Google's product lineup. Gemini 3.8 Live will power Search Live in AI Mode, while Gemini 3.8 Live Extended Thinking is becoming available in Gemini Live.
1
Users with AI Pro and Ultra subscriptions will find Live Extended Thinking in Google Workspace applications including Google Docs, Gmail, and Keep. The model supports recently launched features like Gmail Live for conversational search, Docs Live for draft generation and editing, and Keep Live for note creation.2
The model uses early verbal cues like "Let me check that..." to acknowledge prompts naturally and provides live progress narration to guide users through multi-step background tasks.
2
This approach maintains conversational flow even during complex operations.Related Stories
Developers can access both models through the Gemini API and Google AI Studio starting today.
1
Enterprise users will find both models available in private preview through Gemini Enterprise, with plans to expand availability to Gemini Enterprise for Customer Experience and Google Workspace business customers. All audio generated by these models includes SynthID watermarking, Google's digital watermarking and detection tool that embeds invisible signals into AI-generated content.1
The launch comes approximately two weeks after Google announced Gemini 3.8 Flash and follows the March release of Gemini 3.1 Flash Live.
1
2
These new models demonstrate Google's commitment to near real-time reasoning and making voice interactions with AI feel more natural. For businesses, the combination of strong performance benchmarks and cost efficiency could accelerate adoption of voice agents in customer service and enterprise workflows. Watch for how these models perform in real-world banking and fintech applications, where the τ-Voice-banking benchmark results suggest promising capabilities for handling complex customer interactions.
Source: 9to5Google
Summarized by
Navi
[1]
26 Mar 2026•Technology

12 Dec 2025•Technology

01 Oct 2024

1
Science and Research

2
Policy and Regulation

3
Technology
