GadgetBond

  • Latest
  • How-to
  • Tech
    • AI
      • Apple Intelligence
      • Gemini AI
      • Google DeepMind
      • Anthropic
      • Claude AI
      • Claude Code
      • OpenAI
      • ChatGPT
      • Codex
      • Perplexity
      • SpaceXAI
      • Grok AI
      • Microsoft Copilot
      • Meta AI
    • Amazon
    • Apple
    • CES
    • Computing
    • Creators
    • Google
    • Meta
    • Microsoft
    • Mobile
    • NVIDIA
    • Samsung
    • Security
    • Smart Home
    • Sony
    • Xbox
    • YouTube
  • Transportation
    • Audi
    • BMW
    • Cadillac
    • E-Bike
    • Ferrari
    • Ford
    • Honda Prelude
    • Lamborghini
    • McLaren
    • Mercedes
    • Porsche
    • Rivian
    • Tesla
  • Culture
    • Gaming
    • Streaming
    • Apple TV
    • Disney
    • Hulu
    • Marvel
    • HBO Max
    • Netflix
    • Paramount
    • SHOWTIME
    • Spotify
    • Star Wars
Add GadgetBond as a preferred source to see more of our stories on Google.
Font ResizerAa
GadgetBondGadgetBond
  • Latest
  • Tech
  • AI
  • Deals
  • How-to
  • Apps
  • Computing
  • Gaming
  • Mobile
  • Streaming
  • Transportation
Search
  • Latest
  • Deals
  • Buying Guide
  • How-to
  • Tech
    • Amazon
    • Apple
    • CES
    • Computing
    • Creators
    • Google
    • Meta
    • Microsoft
    • Mobile
    • NVIDIA
    • Samsung
    • Security
    • Smart Home
    • Sony
    • Xbox
    • YouTube
  • AI
    • Apple Intelligence
    • Gemini AI
    • Google DeepMind
    • Anthropic
    • Claude AI
    • Claude Code
    • OpenAI
    • ChatGPT
    • Codex
    • Perplexity
    • SpaceXAI
    • Grok AI
    • Microsoft Copilot
    • Meta AI
  • Transportation
    • Audi
    • BMW
    • Cadillac
    • E-Bike
    • Ferrari
    • Ford
    • Lamborghini
    • McLaren
    • Mercedes
    • Porsche
    • Rivian
    • Tesla
  • Culture
    • Gaming
    • Streaming
    • Apple TV
    • Disney
    • Hulu
    • Marvel
    • HBO Max
    • Netflix
    • Paramount
    • SHOWTIME
    • Spotify
    • Star Wars
Follow US
AIGoogleTech

Google unveils Gemini 3.8 Live models to power next-gen autonomous voice agents

Google’s latest Gemini Live models bring deeper reasoning, visual context and asynchronous tool calling to voice AI.

By
Shubham Sawarkar
Shubham Sawarkar's avatar
ByShubham Sawarkar
Editor-in-Chief
I’m a tech enthusiast who loves exploring gadgets, trends, and innovations. With certifications in CISCO Routing & Switching and Windows Server Administration, I bring a sharp...
Follow:
- Editor-in-Chief
Sep 16, 2026, 2:00 AM EDT
Share
We may get a commission from retail offers. Learn more
Text reading "Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking" centered on a gradient light-blue background, featuring the colorful four-pointed Gemini spark below and blurred white chat bubble icons in the upper right.
Image: Google
SHARE

Google is giving Gemini Live a major upgrade with two new models designed to make real-time AI conversations more capable, natural and useful. The company has introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its latest live dialogue models for voice agents and complex, voice-driven tasks.

Announced on September 15, the new models are built around near-real-time reasoning, allowing Gemini to do more than simply respond to spoken prompts. Gemini 3.8 Live is designed for scalable, cost-efficient conversational experiences, while Gemini 3.8 Live Extended Thinking is aimed at more complicated tasks that require deeper, multi-step reasoning.

The models are also being integrated across Google’s ecosystem, including Gemini Live, Search Live and selected Google Workspace experiences. Developers can access both models through the Gemini API and Google AI Studio.

Gemini can keep talking while it works

One of the biggest changes with Gemini 3.8 Live is the ability to execute tools and API calls in the background without stopping the conversation.

That means a voice agent doesn’t necessarily have to become silent while it waits for an external action to finish. Gemini can acknowledge what the user asked, continue the conversation and handle the underlying task asynchronously.

For example, an agent could start an API request or another tool-based operation while continuing to communicate with the user. Google says this is intended to make interactions feel more like a natural conversation rather than a sequence of prompts followed by long pauses.

Gemini 3.8 Live can also incorporate visual information in near real time. This gives the model additional context about what the user is seeing, opening the door to voice interactions where Gemini can simultaneously listen, look and respond.

Google highlights applications such as real-time employee onboarding and troubleshooting, along with demonstrations involving chess and other visually grounded tasks.

The model can also automatically detect and switch between 97 supported languages during a conversation, making it possible to move between languages without manually changing settings.

Extended Thinking brings deeper reasoning to voice

Gemini 3.8 Live Extended Thinking takes the same basic idea further by adding more intensive reasoning for complex workflows.

Rather than forcing users to wait silently while an AI works through a complicated request, the model can reason in the background while continuing to speak. Google says it can provide early verbal cues such as “Let me check that…” and narrate progress as a multi-step task moves forward.

That combination could be particularly useful for voice agents that need to coordinate several actions. Instead of treating reasoning and conversation as separate phases, the model is designed to handle both simultaneously.

Google describes the model as being intended for enterprise-grade task completion and complex workflows, with support for configurable thinking. Developers can therefore use deeper reasoning when a task requires it rather than treating every voice interaction as a simple question-and-answer exchange.

Google points to strong voice-agent benchmark results

Google is positioning Gemini 3.8 Live Extended Thinking as one of its strongest models for real-time speech-to-speech interaction.

According to Google’s reported benchmarks, the model scored 82.6 on Artificial Analysis’ Speech to Speech Quality Index, where Google says it took the overall top position. It also reported scores of 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark for agentic task completion. On Big Bench Audio, the model scored 97.7%.

Gemini 3.8 Live, meanwhile, placed second in Google’s cited Speech Agent Arena results.

These are Google’s reported benchmark results, so they should be viewed in the context of the particular tests, configurations and evaluation environments used. Google also says both models performed well on ServiceNow’s EVA-Bench, where the company evaluated the balance between task completion and conversational quality.

Built for developers and voice agents

The new models aren’t limited to Google’s own consumer applications. Google is also making them available as part of its developer-focused Gemini Live API.

Developers can use Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to build voice-driven applications that combine speech, reasoning, visual context and tool execution.

Google lists asynchronous function calling, visual context, multilingual support, alphanumeric precision and incremental content updates among the capabilities available to developers. The company says the models can handle things such as confirmation codes, claim numbers and other technical information where accurate handling of spoken alphanumeric data matters.

The models are also supported through a growing collection of development platforms, including Agora, Fishjam, LiveKit, LangChain, Pipecat, Vercel and Vision Agents. These platforms can handle parts of the real-time media infrastructure, allowing developers to focus more heavily on the actual voice-agent experience.

For developers using the Live API, Google lists pricing of $0.005 per minute for audio input and $0.018 per minute for audio output.

Gemini 3.8 Live is coming to more Google products

The new models are also being pushed into Google’s own products.

Gemini 3.8 Live is rolling out to developers through the Gemini API and Google AI Studio, while enterprise customers can access it through a private preview in Gemini Enterprise. Google says it is also available to everyone through Search Live.

Gemini 3.8 Live Extended Thinking is similarly rolling out through the Gemini API and Google AI Studio. On the consumer side, Google says it is available in Gemini Live, while Google AI Pro and Ultra subscribers can use it in Docs through Workspace. Gmail and Keep are also getting the model for Google AI subscribers.

Google is effectively turning voice interaction into another interface for getting work done rather than treating it purely as a conversational feature. The combination of background tool execution, visual understanding and simultaneous reasoning could make that distinction increasingly important.

Google is also watermarking AI-generated audio

Google says audio generated by its AI products includes SynthID, the company’s imperceptible watermarking technology.

The watermark is embedded into generated audio so that AI-created content can be identified, according to Google. The company says this is intended to improve transparency and help address the potential spread of misinformation involving synthetic audio.

That becomes increasingly relevant as voice models become more convincing and capable of carrying on longer, more natural conversations.

With Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, Google is moving its live AI models toward a model of interaction where the assistant doesn’t simply wait for a prompt, generate a response and stop. Instead, it can listen, see, reason, talk and work on tasks at the same time.

For users, that could make Gemini Live feel considerably less like a voice chatbot and more like an always-available voice interface for getting things done.


Discover more from GadgetBond

Subscribe to get the latest posts sent to your email.

Topic:Gemini AI (formerly Bard)
Leave a Comment

Leave a ReplyCancel reply

Most Popular

Apple starts 27.2 beta cycle with iPhone Duo support coming
Amazon Prime Big Deal Days returns October 6–7 for 48-hour fall shopping event
Canva launches ProSuite with Affinity, Cavalry, Flourish and Leonardo
Sony reveals next-generation Pulse wireless headsets for PS5
Porsche Cayenne Electric adds wireless charging and electric doors

Also Read

A row of popular non-fiction ebook covers floating above an AI prompt bar reading "Help me personalize learnings from this book". The featured book covers include Thinking in Bets, The Sense of Style, The Lean Startup, Radical Candor, The New Menopause, Food Rules, and The Power of Habit.

Google launches Expert Intelligence for Gemini Notebook

A promotional graphic set against a soft blue and lavender gradient background. On the left is the white text "CC" next to an outlined pill badge reading "EXPERIMENT." On the right, layered digital interface cards tilt forward into view: a central card titled "Your Day Ahead" details a personalized family agenda with sections like "Top of mind," "On your calendar" (featuring drop-off, pickup, and appointment times), and "On your list." In the background sits a Google Calendar view with an open event card titled "[Maya] Swim Class Youth Level 2," displaying activity details, dates, and instructions.

Google’s CC AI agent can now manage family calendars, tasks, and chores

Chris Pratt as Navy SEAL Commander James Reece in The Terminal List Season 2, wearing tactical military gear and a body armor vest, looking alertly to his left against a dim, out-of-focus background.

Prime Video sets October 21 premiere for The Terminal List Season 2

Mozilla logo

Mila, Mozilla, and Hypertec partner to build local, sovereign AI in Canada

ASUS Ascent QN10 hero

ASUS unveils Ascent QN10 with Snapdragon X2 Elite

Three Apple iPhone SE 3 smartphones displayed diagonally at an angle against a white background. Each device showcases a front screen with prominent top and bottom bezels and a Touch ID home button, featuring vibrant multi-colored gradient wallpapers with horizontal light beam patterns. The phones highlight the three color finishes: a red aluminum frame at the bottom, a dark midnight frame in the center, and a pale starlight frame at the top

I installed iOS 27 on my iPhone SE 3. Here’s everything that’s new—or, in some cases, what’s not

A collage of user interface screens showcasing new podcast features on the Threads app against a black background, including analytics insights, post creation tools, video episode previews, and interactive transcript cards.

Threads launches new tools for podcast creators

Front view of glossy black Snap SPECS augmented reality glasses featuring clear rectangular lenses, built-in side camera sensors, and the "SPECS" wordmark centered below against a plain white background.

Snap launches $2,195 SPECS AR glasses and SPECS Intelligence

Company Info
  • Homepage
  • Support my work
  • Latest stories
  • Company updates
  • GDB Recommends
  • Daily newsletters
  • About us
  • Contact us
  • Write for us
  • Editorial guidelines
Legal
  • Privacy Policy
  • Cookies Policy
  • Terms & Conditions
  • DMCA
  • Disclaimer
  • Accessibility Policy
  • Security Policy
  • Do Not Sell or Share My Personal Information
Socials
Follow US

Disclosure: We love the products we feature and hope you’ll love them too. If you purchase through a link on our site, we may receive compensation at no additional cost to you. Read our ethics statement. Please note that pricing and availability are subject to change.

Copyright © 2026 GadgetBond. All Rights Reserved. Use of this site constitutes acceptance of our Terms of Use and Privacy Policy | Do Not Sell/Share My Personal Information.