Google Launches Gemini 3.5 Live Translate for Real-Time Speech Interpretation in More Than 70 Languages
Google is extending Gemini 3.5's generative AI capabilities to real-time interpretation through direct speech-to-speech translation that preserves a speaker's original tone, cadence and delivery. Compared with conventional systems that first convert speech into text before synthesizing audio, the new model can reduce pauses, making it significant for international meetings, travel and customer service communications.
Google has launched the Gemini 3.5 Live Translate audio model, which can translate more than 70 languages in real time. The model is being integrated into Google Translate and Google Meet and is also available in developer preview. Google has begun testing it with partners including Southeast Asian ride-hailing and delivery platform Grab to assess translation speed and naturalness in real-world service settings.
All Coverage
4 original reportsThe Backstory
The history behind this eventGoogle Launches Gemini 3.5 Transcribe to Filter Fillers and Noise
Google is expanding its Gemini audio lineup for transcription, meeting notes, captions, customer service and voice-controlled applications. Speech-to-text systems often lose accuracy when speakers overlap, conversations are interrupted or background noise obscures words. Gemini 3.5 Transcribe is designed to use context to recognize specialized terminology and produce cleaner prose, reducing the manual editing typically required after automated transcription.
The new model supports more than 85 languages and recorded an average word error rate of 2.6%, according to Google. It can distinguish as many as three speakers, remove fillers such as “ums” and “ahs,” and provide real-time translation while handling noisy or interrupted speech. Traditional Chinese is not currently listed among the supported languages, limiting its immediate usefulness for some Taiwan-based users.
Google Translate Marks 20 Years With Gemini-Powered Live Translation and AI Conversation Practice
Google Translate grew out of a machine-learning experiment launched by Google Research in 2006. It shifted from statistical methods to neural networks in 2016 and now supports nearly 250 languages, translating 1 trillion words each month. The service has moved cross-language communication beyond word-for-word conversion toward contextual understanding, reaching more than 1 billion users worldwide.
Google announced its 20th-anniversary updates on April 28, 2026. The Android app added AI pronunciation practice, initially supporting English, Spanish and Hindi in the United States and India. The feature can analyze speech and provide real-time feedback. Gemini models have also improved the translation of idioms and local slang, as well as live translation. Personalized conversation practice offers listening and speaking lessons built around real-world scenarios and tailored to each learner’s proficiency level.
Google Launches Gemini 3.1 Flash TTS With Support for 70 Languages and Scene Direction
Google is extending Gemini’s multimodal capabilities into speech generation, enabling developers to create more natural interactions for customer service, video dubbing and accessibility services. Gemini 3.1 Flash TTS supports multiple languages and emotional control, lowering the barrier to producing voice content across markets and making it particularly important for localizing AI applications.
As of July 19, 2026, Google AI developer relations lead Logan Kilpatrick announced the launch of Gemini 3.1 Flash TTS. The model supports more than 70 languages and lets developers use audio tags to direct scenes and adjust emotion and delivery. It is now available to try in Google AI Studio and can be integrated through the Gemini API, though pricing has not yet been announced.
Google Launches Gemini 3.1 Flash Live, Expands Search Live to More Than 200 Countries
Google first launched Search Live in the United States in September 2025, extending generative AI beyond text search to continuous voice- and camera-based follow-up questions. The service combines Google’s search index with AI Mode, breaking down questions during conversations and providing links to webpages. The move reflects Google’s effort to reshape its core search gateway with Gemini. The company did not disclose the investment behind the latest rollout.
On March 26, 2026, Google released Gemini 3.1 Flash Live, a real-time audio model designed to improve the speed, naturalness and stability of multilingual responses. It also expanded Search Live to more than 200 countries and territories where AI Mode is available, with support for 98 languages. Users can tap Live in the Google app on Android or iOS, or use Google Lens to conduct continuous searches by combining the camera with voice queries.
Subscribe to Mark Radar Weekly
Every Friday, the week's strongest signals in your inbox. Unsubscribe anytime.
If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →