GadgetBond

  • Latest
  • How-to
  • Tech
    • AI
      • Apple Intelligence
      • Gemini AI
      • Google DeepMind
      • Anthropic
      • Claude AI
      • Claude Code
      • OpenAI
      • ChatGPT
      • Codex
      • Perplexity
      • SpaceXAI
      • Grok AI
      • Microsoft Copilot
      • Meta AI
    • Amazon
    • Apple
    • CES
    • Computing
    • Creators
    • Google
    • Meta
    • Microsoft
    • Mobile
    • NVIDIA
    • Samsung
    • Security
    • Smart Home
    • Sony
    • Xbox
    • YouTube
  • Transportation
    • Audi
    • BMW
    • Cadillac
    • E-Bike
    • Ferrari
    • Ford
    • Honda Prelude
    • Lamborghini
    • McLaren
    • Mercedes
    • Porsche
    • Rivian
    • Tesla
  • Culture
    • Gaming
    • Streaming
    • Apple TV
    • Disney
    • Hulu
    • Marvel
    • HBO Max
    • Netflix
    • Paramount
    • SHOWTIME
    • Spotify
    • Star Wars
Add GadgetBond as a preferred source to see more of our stories on Google.
Font ResizerAa
GadgetBondGadgetBond
  • Latest
  • Tech
  • AI
  • Deals
  • How-to
  • Apps
  • Computing
  • Gaming
  • Mobile
  • Streaming
  • Transportation
Search
  • Latest
  • Deals
  • Buying Guide
  • How-to
  • Tech
    • Amazon
    • Apple
    • CES
    • Computing
    • Creators
    • Google
    • Meta
    • Microsoft
    • Mobile
    • NVIDIA
    • Samsung
    • Security
    • Smart Home
    • Sony
    • Xbox
    • YouTube
  • AI
    • Apple Intelligence
    • Gemini AI
    • Google DeepMind
    • Anthropic
    • Claude AI
    • Claude Code
    • OpenAI
    • ChatGPT
    • Codex
    • Perplexity
    • SpaceXAI
    • Grok AI
    • Microsoft Copilot
    • Meta AI
  • Transportation
    • Audi
    • BMW
    • Cadillac
    • E-Bike
    • Ferrari
    • Ford
    • Lamborghini
    • McLaren
    • Mercedes
    • Porsche
    • Rivian
    • Tesla
  • Culture
    • Gaming
    • Streaming
    • Apple TV
    • Disney
    • Hulu
    • Marvel
    • HBO Max
    • Netflix
    • Paramount
    • SHOWTIME
    • Spotify
    • Star Wars
Follow US
AIGoogleTech

Google’s Gemini 3.5 Live Translate brings natural speech translation to life

The new audio model preserves intonation, pacing, and pitch—qualities most AI voice tools strip away in favor of robotic flatness.

By
Shubham Sawarkar
Shubham Sawarkar's avatar
ByShubham Sawarkar
Editor-in-Chief
I’m a tech enthusiast who loves exploring gadgets, trends, and innovations. With certifications in CISCO Routing & Switching and Windows Server Administration, I bring a sharp...
Follow:
- Editor-in-Chief
Jun 10, 2026, 9:00 AM EDT
Share
We may get a commission from retail offers. Learn more
Promotional graphic for Gemini 3.5 Live Translate featuring Google's multicolored Gemini logo alongside the text “Gemini 3.5 Live Translate” centered on a soft white and blue abstract gradient background. The minimalist design represents Google's AI-powered real-time translation capabilities for live multilingual conversations and language assistance.
Image: Google
SHARE

Twenty years ago, Google started one of its first machine learning experiments with a pretty simple goal: turn the science of language into the magic of human connection. That experiment became Google Translate, and today, the company translates over a trillion words every month for billions of users across its products. But on June 9, 2026, Google announced it’s taking that experiment to the next level with Gemini 3.5 Live Translate, its latest audio model designed for live, real-time speech-to-speech translation.

What makes this different from everything we’ve seen before? Well, for starters, it actually works at the speed of human conversation. Unlike the turn-by-turn translation systems that wait for you to finish speaking before responding, Gemini 3.5 Live Translate generates speech continuously. It stays just a few seconds behind the speaker throughout the entire session, delivering fluid audio without awkward pauses. Google says it’s fast enough to keep up with a normal conversation, which is a pretty bold claim when you think about how most translation tools still feel like they’re running through a dial-up connection.

The model automatically detects more than 70 languages and generates smooth, natural-sounding translated speech that preserves the speaker’s intonation, pacing, and pitch. This is a big deal because so many AI voice models sound robotic or flat, stripping away the personality and emotion from what someone’s actually saying. But Gemini 3.5 Live Translate keeps those human qualities intact, which makes the translation feel less like you’re talking to a machine and more like you’re having a real conversation with someone who just happens to speak a different language.

If you’re wondering how this actually works under the hood, the model processes speech as it’s streamed, enabling a more seamless connection across languages. It handles multilingual inputs without requiring you to manually configure any settings, and its noise robustness means it can function in loud, unpredictable environments. So if you’re trying to translate a conversation at a busy airport or a crowded street market, it won’t just break down because of background noise.

Google is rolling out Gemini 3.5 Live Translate across three different surfaces starting today. Developers get it in public preview through the Gemini Live API and Google AI Studio, which means they can start building voice translation apps into their own platforms right now. Enterprises are getting a private preview in Google Meet starting this month, and everyone else can use it through the Google Translate app on both Android and iOS.

For developers, the integration is pretty straightforward. The Gemini Live API supports low-latency, real-time speech-to-speech translation between 70+ languages using the gemini-3.5-live-translate-preview model. By configuring the API with translation settings, you can stream audio in one language and receive translated audio output in another, enabling seamless real-time voice-to-voice translation. Developer platforms like Agora, Fishjam, LiveKit, Pipecat, and Vision Agents are already integrating the technology to enable voice translation applications, which means they’re handling the complex real-time media streaming infrastructure so developers can focus on the user experience.

One of the companies already testing this is Grab, the Southeast Asian tech giant. They’re using Gemini 3.5 Live Translate to enable multilingual communication in near real-time between drivers and travelers at pickups. These users make over 10 million voice calls per month through Grab, so having a translation tool that actually works in real-time is going to be a huge improvement for their service. Philipp Kandal, Grab’s Chief Product Officer, said they’ve valued the model’s ability to auto-detect multiple languages and translate speech accurately with low latency.

In Google Meet, speech translation is going to get a major upgrade. The previous limit was just five languages, but with Gemini 3.5 Live Translate, it’s expanding to over 70 languages. That means conversations across more than 2,000 language combinations in one meeting, which is a massive jump from the previous state of only translating to and from English. The interface is also getting updated to provide instant access to speech translation, so you won’t have to dig through settings menus to find it. Google is launching this in private preview for select business Google Workspace customers starting this month, followed by a broader rollout later in the year.

For regular users on the Google Translate app, the experience is pretty slick. When using the Live translate feature, you just connect any pair of headphones and experience more seamless translation that mirrors the speaker’s tone across 70+ languages. Android users are also getting a new “listening mode” that lets you hear translations directly through your phone’s earpiece. You hold your phone to your ear like a regular call, and the translated audio streams straight to you. This is helpful when you want to quickly hear translations without others hearing and you don’t have headphones handy.

There’s also a safety consideration here that Google is being upfront about. All audio generated by Gemini 3.5 Live Translate is watermarked with SynthID, an imperceptible watermark woven directly into the audio output. This ensures AI-generated content remains detectable to help prevent misinformation, which is becoming increasingly important as AI voice technology gets more advanced.

What’s really interesting about this release is how it represents a shift in what we expect from translation technology. For years, we’ve been stuck with tools that work, but they don’t feel natural. There’s always a delay, a robotic quality, or a sense that you’re not actually having a conversation but rather exchanging pre-programmed messages. Gemini 3.5 Live Translate is trying to close that gap, and based on the early feedback from companies like Grab, CJ ENM, and LiveKit, it’s actually succeeding.

Tech journalists and developers who’ve gotten early access have shared positive feedback highlighting the impressive translation quality, accuracy, and low latency. The model’s ability to auto-detect languages without manual configuration is a standout feature, and the continuous stream translation approach means it doesn’t have to wait until one person has finished speaking before it starts generating a response.

Looking at the bigger picture, this is part of Google’s broader push into AI audio models. Earlier this year, the company introduced Gemini Omni and other 3.5 models that showed off computer use capabilities and advanced audio processing. Gemini 3.5 Live Translate fits into that ecosystem as a specialized tool for one of the most practical applications of AI: making it possible for people who speak different languages to communicate naturally.

The timing is also significant. As global communication becomes more important in business, travel, and everyday life, the ability to translate conversations in real-time is becoming less of a luxury and more of a necessity. Whether you’re a developer building a communication app, a business running international meetings, or just someone trying to navigate a foreign country, this kind of technology has real-world value.


Discover more from GadgetBond

Subscribe to get the latest posts sent to your email.

Leave a Comment

Leave a ReplyCancel reply

Most Popular

Apple removes iPhone 17 Pro models from its store
Apple unveils new iPhone 18 Pro cases and wrist strap
Apple adds AI-powered health features to Apple Watch and iPhone
iPhone 18 Pro introduces Apple’s first variable-aperture camera
Here’s everything Apple announced at its September 9 event

Also Read

Four Ring Indoor Cameras wearing NFL team-themed helmet covers for the Jets, Seahawks, Patriots, and Rams.

Ring made an NFL helmet for your indoor security camera

A horizontal side profile of a burgundy iPhone 18 Pro lying flat against a black background, set in front of large, reflective metallic text spelling "PRO" with subtle iridescent light refractions.

AT&T unveils iPhone 18 Pro, Pro Max and Duo offers

Apple iPhone Duo shown folded and unfolded to compare its outer and larger inner displays

Verizon reveals iPhone 18 Pro, Apple Watch and AirPods offers

The new BMW 3 Series camouflaged prototypes in Miramas.

BMW confirms new 3 Series with four- and six-cylinder engines

A promotional graphic for Adobe Acrobat showcasing an "Organic Chemistry Student Space." Chemistry PDF documents and notes are uploaded on the left, flanked by student profile avatars. In the center, yellow, green, and blue glass laboratory flasks hold colorful flowers, set against a bright yellow background. On the right, a white menu displays options to "Create" a "Study Guide," "Practice Quiz," or "Flashcards," with a cursor pointing toward "Flashcards."

Adobe Acrobat Student Spaces is now free worldwide

A screenshot of Adobe Premiere’s editing timeline featuring the Generative Media interface. A video clip on the timeline displays a woman wearing sunglasses and a yellow dress outdoors by a poolside table. An eyedropper cursor samples this clip as the "First frame" reference. The generative tool popup shows a text prompt starting with "Rising pull-back revealing neighborho…", with settings set to the Kling AI model, 1080p resolution, and 16:9 aspect ratio. Audio waveforms in green run beneath the video tracks.

Adobe unveils new AI-powered tools for Premiere and After Effects

A 3D graphic of digital document cards set against a warm pink and yellow gradient background. The central card displays a dark background with vibrant, glowing purple and blue floral petals, overlaid with white text that reads "Master services agreement." Surrounding the card are three white floating UI buttons with icons that say "Filter documents," "Analyze files in bulk," and "Export data to report."

Adobe Acrobat Studio can now search documents and analyze contracts

A geometric flat-art illustration centered on a dark green background, depicting security and data protection motifs. It features an arrangement of black and pastel-toned rectangular blocks, diagonal purple-and-black hatched patterns, and stylized gold keys and keyholes framing a large concentric circular lock mechanism.

Figma enterprise files can now be hosted in Japan

Company Info
  • Homepage
  • Support my work
  • Latest stories
  • Company updates
  • GDB Recommends
  • Daily newsletters
  • About us
  • Contact us
  • Write for us
  • Editorial guidelines
Legal
  • Privacy Policy
  • Cookies Policy
  • Terms & Conditions
  • DMCA
  • Disclaimer
  • Accessibility Policy
  • Security Policy
  • Do Not Sell or Share My Personal Information
Socials
Follow US

Disclosure: We love the products we feature and hope you’ll love them too. If you purchase through a link on our site, we may receive compensation at no additional cost to you. Read our ethics statement. Please note that pricing and availability are subject to change.

Copyright © 2026 GadgetBond. All Rights Reserved. Use of this site constitutes acceptance of our Terms of Use and Privacy Policy | Do Not Sell/Share My Personal Information.