ViqusViqus
Navigate
Company
Blog
About Us
Contact
System Status
Enter Viqus Hub

Google Launches Gemini 3.8 Live: Advances Voice AI with Near Real-Time Reasoning and Tool Execution

Gemini 3.8 Live real-time reasoning voice-based artificial intelligence tool calls Google AI Speech-to-Speech Quality Index
September 15, 2026
Viqus Verdict Logo Viqus Verdict Logo 7
High-Fidelity Conversation Meets Asynchronous Computing
Media Hype 6/10
Real Impact 7/10

Article Summary

Google announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, positioning them as leading advancements in voice-based AI agents. The primary focus is mitigating the perceived latency of conversational AI by enabling near-real-time reasoning and simultaneous speech-and-thought processing. Critically, these models can execute third-party tool calls and APIs in the background without disrupting the natural flow of conversation, offering a more seamless user experience. Beyond real-time capabilities, the models boast advanced features including automatic language detection (supporting 97 languages), near real-time visual grounding, and early verbal cue recognition, making interactions feel more human-like. The models are available via APIs and Google Workspace, with clear pricing tiers outlined for both audio inputs and outputs.

Key Points

  • The new Gemini 3.8 Live models specialize in managing background task execution (tool calls, API integrations) while maintaining a natural, uninterrupted conversational pace.
  • Technological improvements include near real-time reasoning, multi-language support (97 languages), and advanced conversational cues, enhancing the human-like feel of AI interaction.
  • Google has positioned the tools for enterprise adoption, making them available through the Gemini API, Google Workspace, and emphasizing developer integration via partner platforms.

Why It Matters

This is a significant iteration in the race for practical, commercial AI agents. The breakthrough isn't just the improved accuracy or benchmark score; it's the ability to convincingly blend complex, asynchronous background computation (e.g., booking a trip, searching the web) with natural, uninterrupted dialogue. For professionals, this means the next generation of AI assistants will move past simply providing information and will become reliable, hands-off operational workers, fundamentally changing the design requirements for enterprise SaaS tools and customer service interactions.

You might also be interested in