Ted Hisokawa Sep 15, 2026 19:07

Google introduces Gemini 3.8 Live and Extended Thinking, advanced AI models for real-time voice agents and complex task reasoning.

Google Launches Gemini 3.8 Live Models for Real-Time AI Dialogue

Google has unveiled its latest AI advancements with the release of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two real-time dialogue models designed to enhance voice agent capabilities. Announced on September 15, 2026, these models aim to redefine user interaction with AI by improving conversational fluidity, task complexity handling, and response speed.

The Gemini 3.8 Live model is optimized for low-latency interactions, making it ideal for fast-paced voice agents and real-time dialogue. In contrast, Gemini 3.8 Live Extended Thinking is tailored for high-complexity tasks requiring multi-step reasoning, such as technical diagnostics or multi-API travel bookings. Both are built on Google’s Gemini 3 platform, which emphasizes multimodal reasoning, cost efficiency, and scalability.

Performance Benchmarks and Capabilities

Gemini 3.8 Live Extended Thinking has secured top rankings on several industry benchmarks. It achieved the highest score on Artificial Analysis’ Speech to Speech Quality Index (82.6) and excelled in agentic task completion benchmarks, including a 68.6% performance on τ-Voice and a 97.7% score on Big Bench Audio reasoning tasks. These metrics signal a significant leap in AI’s ability to handle complex workflows while maintaining conversational flow.

Meanwhile, Gemini 3.8 Live demonstrated its efficiency by ranking second in Speech Agent Arena benchmarks, balancing performance and cost for developers building scalable voice-driven applications. Both models support 97 languages and seamlessly integrate visual context, enabling AI to process and respond to visual inputs in near real-time.

Use Cases for Developers and Enterprises

For developers, Gemini 3.8 models are available through the Gemini API and Google AI Studio. These tools enable the creation of voice-driven applications with real-time reasoning, background task execution, and conversational multitasking. Integration partners such as Agora, LangChain, and LiveKit are already leveraging these capabilities to streamline deployment for high-performance voice agents.

Enterprises can preview the models via the Gemini Enterprise platform, with broader access coming to Google Workspace and Search Live. Use cases span technical support, employee onboarding, and even sophisticated tasks like STEM tutoring and business planning, where conversational AI can narrate progress and offer real-time updates.

Focus on Safety and Transparency

All audio generated by Gemini 3.8 models includes SynthID watermarking, an imperceptible marker ensuring that AI-generated content remains identifiable. This aligns with Google’s broader commitment to responsible AI deployment and mitigating risks like misinformation.

Rollout and Availability

Both models are rolling out starting today. Developers can access them via the Gemini API, while enterprises can explore private previews through Google’s Vertex AI Studio. Integration with Google Workspace tools like Docs Live and Gmail Live will follow in the coming months, bringing these capabilities to a wider audience.

With Gemini 3.8 Live and Extended Thinking, Google continues to push the boundaries of real-time AI dialogue, offering tools that are not only intelligent but also practical and cost-effective for developers and enterprises alike.

Image source: Shutterstock Source

LEAVE A REPLY

Please enter your comment!
Please enter your name here