Skip to content
Break Read Break Read Break Read
Break Read Break Read Break Read
  • Blog
  • Contact
  • Blog
  • Contact
Close

Search

Home/AI/Google Launches Gemini 3.8 Live With Real-Time Voice and Background Reasoning
Gemini 3.8 Live
AI

Google Launches Gemini 3.8 Live With Real-Time Voice and Background Reasoning

September 16, 2026 4 Min Read

Table of Contents

What Is Gemini 3.8 Live?
Extended Thinking Adds Deeper Reasoning
Gemini 3.8 Live Supports Visual Context
97 Languages and Multilingual Conversations
Gemini 3.8 Live Pricing
Gemini 3.8 Live Scores 82.6 on Artificial Analysis
Gemini 3.8 Live vs. OpenAI Voice AI
Developer and Enterprise Availability
What Gemini 3.8 Live Means for Voice AI
FAQS
Conclusion

Google has launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two native speech-to-speech AI models designed for real-time voice applications and natural conversations.

Announced on September 15, 2026, the models are aimed at developers building voice agents that can understand speech, process visual information, use external tools and continue conversations while tasks run in the background. Google describes them as its most advanced live dialogue models yet.

What Is Gemini 3.8 Live?

Gemini 3.8 Live is designed for low-latency voice interactions. It supports asynchronous function calling, allowing external tools and API requests to run while the conversation continues.

This could support voice agents handling customer service, scheduling, research and other workflows requiring external services.

FeatureGemini 3.8 Live
Model typeNative speech-to-speech
Primary useReal-time voice agents
Visual inputSupported
Tool callingSupported
Asynchronous functionsSupported
Language support97+ languages
Developer accessGemini API and Google AI Studio
LaunchSeptember 15, 2026

Extended Thinking Adds Deeper Reasoning

Gemini 3.8 Live Extended Thinking is designed for more complex, multi-step tasks.

Google says the model can reason while continuing to speak and perform background tool calls without completely stopping the conversation.

For example, a voice agent could acknowledge a request, perform several external actions and return with the completed result.

This makes Extended Thinking relevant to agentic applications combining conversation, reasoning and external actions.

Gemini 3.8 Live Supports Visual Context

The models can process visual information alongside voice.

Users can provide images, video or other visual context while interacting with the AI. Potential applications include visual assistance, learning, technical support and software development.

Google’s demonstrations include visual interactions, coding workflows and multi-step tasks. These demonstrate potential capabilities rather than guaranteed performance across every application.

97 Languages and Multilingual Conversations

Google says Gemini 3.8 Live supports 97+ languages and can switch between supported languages during conversations.

This could benefit international voice applications, including customer-service tools that handle multilingual interactions.

Gemini 3.8 Live Pricing

Google’s Gemini API pricing lists $0.005 per minute for audio input and $0.018 per minute for audio output on the standard paid tier.

Using those rates for one hour of audio input and one hour of audio output produces an illustrative cost of approximately $1.38.

However, this is not a fixed hourly price. Actual costs can vary based on token usage, conversation context and other API features. Therefore, claims that Gemini is a specific number of times cheaper than competing voice models can be misleading without comparing identical workloads.

Gemini 3.8 Live Scores 82.6 on Artificial Analysis

Google says Gemini 3.8 Live Extended Thinking scored 82.6 on Artificial Analysis’ Speech-to-Speech Quality Index, while Gemini 3.8 Live scored 76.0.

Google also reports 68.6% on the τ-Voice benchmark and 97.7% on Big Bench Audio for Extended Thinking.

These figures measure specific capabilities and should not be treated as a universal ranking of voice-AI quality. Performance can vary by workload, latency requirements, tools and application design.

Gemini 3.8 Live vs. OpenAI Voice AI

The launch brings Google into closer competition with OpenAI and other companies developing real-time voice agents.

Developers comparing these systems need to consider audio quality, latency, reasoning, tool execution, pricing, context handling and infrastructure rather than relying on a single benchmark.

Developer and Enterprise Availability

Google has made the models available through the Gemini API and Google AI Studio.

Gemini 3.8 Live is also being used in Google’s Search Live experience, while Extended Thinking is being introduced across Gemini Live and selected Google Workspace experiences.

Google has highlighted partnerships with companies and platforms including Agora, LiveKit, Pipecat, Vercel and Salesforce.

Google also says audio generated by the models is watermarked with SynthID.

What Gemini 3.8 Live Means for Voice AI

Voice agents are increasingly expected to do more than recognize speech and answer questions. They can now combine visual context, reasoning, external tools and ongoing conversation.

Gemini 3.8 Live Extended Thinking brings these capabilities together with background reasoning, moving voice AI toward systems that can talk, reason and act within the same interaction.

FAQS

What is Gemini 3.8 Live?

Gemini 3.8 Live is Google’s native speech-to-speech model designed for low-latency voice conversations and voice-agent applications.

What is Gemini 3.8 Live Extended Thinking?

It is designed for complex tasks, adding deeper reasoning and background task execution to live voice interactions.

How much does Gemini 3.8 Live cost?

Google lists audio input at $0.005 per minute and audio output at $0.018 per minute. Actual costs depend on usage and workload.

Can Gemini 3.8 Live use external tools?

Yes. The Live API supports function calling and asynchronous tool execution.

Where can developers access Gemini 3.8 Live?

Developers can access the models through the Gemini API and Google AI Studio.

Conclusion

Google’s launch of Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking expands its real-time voice AI capabilities.

The models combine speech-to-speech interaction with visual understanding, asynchronous tool calling and deeper reasoning. Google’s pricing provides another option for developers, although actual costs depend on usage.

The reported benchmark results offer measurements of specific capabilities rather than an overall ranking.

The larger development is the ability of voice agents to converse while reasoning and taking action, a combination that could become increasingly important for AI applications.

Share this article on
  • Facebook
  • Pinterest
  • Twitter
  • Linkedin
  • Whatsapp
Author

Lalith Raj

Follow Me
Other Articles
Chandrayaan-1
Previous

Chandrayaan-1 Data Reveal Evidence of Ancient Australe Basin on Moon

Asian Games 2026
Next

Asian Games 2026 Begin in Aichi-Nagoya With 43 Sports

Search...

Recent Posts

  • Asian Games 2026
    Asian Games 2026 Begin in Aichi-Nagoya With 43 Sports
    by Lalith Raj
    September 16, 2026
  • Snapchat just brought AI powered conversational ads to its app. 2
    Snapchat Launches Sponsored Interactive AI Ads Inside Chat
    by Nithin
    March 1, 2026
  • Lovable just launched its vibe coding app on iOS and Android
    Lovable Mobile App Launches Vibe Coding Experience on iOS and Android
    by Nithin
    March 4, 2026
  • Apple just introduced a cheaper option for App Store subscriptions
    Apple introduces a new subscription model: Monthly Plans with 12-Month Commitment
    by Nithin
    March 8, 2026

Categories

  • AI
  • Business
  • Cars
  • Entertainment
  • Finance
  • Music
  • News
  • Science
  • SEO
  • Sports
  • Technology
  • Trending

Break Read

Stay ahead in the fast-moving world of technology with expert articles, industry updates, and practical insights.

Latest Posts

  • Asian Games 2026 Begin in Aichi-Nagoya With 43 SportsSeptember 16, 2026
  • Google Launches Gemini 3.8 Live With Real-Time Voice and Background ReasoningSeptember 16, 2026
  • Chandrayaan-1 Data Reveal Evidence of Ancient Australe Basin on MoonSeptember 16, 2026

Pages

  • Contact
  • Terms and Conditions
  • Privacy Policy
  • Refund Policy
Copyright 2026 — Break Read. All rights reserved.
Go to mobile version