Google DeepMind Releases Gemini 3.5 Live Translate for Real-Time Speech Translation Across 70+ Languages
Google DeepMind released Gemini 3.5 Live Translate, an audio model that provides near real-time speech-to-speech translation across 70+ languages. The model automatically detects languages, preserves speaker intonation and pacing, and maintains a few seconds of latency while generating continuous speech output.
Google DeepMind Releases Gemini 3.5 Live Translate for Real-Time Speech Translation Across 70+ Languages
Google DeepMind released Gemini 3.5 Live Translate on June 9, 2026, an audio model that provides near real-time speech-to-speech translation across 70+ languages with automatic language detection.
Technical Capabilities
The model generates continuous translated speech while maintaining a latency of "just a few seconds" behind the speaker, according to Google. Unlike turn-based translation systems that wait for complete sentences, Gemini 3.5 Live Translate processes streaming audio and balances translation speed with contextual accuracy.
Key technical features include:
- Automatic detection of 70+ languages without manual configuration
- Preservation of speaker intonation, pacing, and pitch in translated output
- Noise robustness for unpredictable environments
- Support for over 2,000 language pair combinations in single sessions
- SynthID watermarking embedded in all generated audio
Availability and Deployment
Gemini 3.5 Live Translate is rolling out across three channels:
Gemini Live API: Available in public preview for developers via Google AI Studio. Developer platforms including Agora, Fishjam, LiveKit, Pipecat, and Vision Agents have integrated the API for real-time media streaming infrastructure.
Google Meet: Launching in private preview this month for select Google Workspace business customers, expanding from the previous limitation of five languages and English-only translation pairs. Broader rollout planned for later in 2026.
Google Translate app: Rolling out globally on Android and iOS. The model powers the Live translate feature for users with connected headphones. Android users receive an additional "listening mode" that streams translations through the phone's earpiece without headphones.
Early Implementations
Grab, which processes over 10 million voice calls monthly, is testing the model to enable multilingual communication between drivers and travelers. Additional partners including CJ ENM and LiveKit have provided feedback on translation quality and low latency, according to Google.
Pricing for API access has not been disclosed.
What This Means
Gemini 3.5 Live Translate represents Google's entry into the competitive real-time speech translation market, directly challenging established players in multilingual communication tools. The 70+ language support and 2,000+ language pair combinations significantly exceed the capabilities of Google's previous Meet translation system, which supported only five languages with English as a required pivot.
The model's continuous streaming approach addresses a core limitation of turn-based systems, though the "few seconds" latency specification lacks precision for developers evaluating real-time requirements. The integration across Google's product ecosystem—from developer APIs to consumer apps—indicates a platform play rather than a standalone model release. However, the lack of disclosed API pricing and benchmark comparisons to competing speech translation models limits technical evaluation.
Related Articles
Google Launches Lyria 3.5 AI Music Model Directly Inside the Gemini App
Google has released Lyria 3.5, a new AI music generation model, directly inside the Gemini app alongside availability in AI Studio, Flow Music, and Vids. Google claims the model was trained exclusively on licensed content and produces more expressive vocals than its predecessor.
Google Launches WeatherNext 3, Claims 50% More Accurate Precipitation Forecasts
Google DeepMind and Google Research released WeatherNext 3, a weather AI model trained on real-time geostationary satellite data instead of lagging numerical weather prediction outputs. Google claims up to 50% more accurate day-ahead precipitation forecasts, now rolling out to Search, Maps, and the Gemini app.
Google Launches Gemini 3.8 Flash With Same Pricing as Predecessor, Its Third Flash Model in Six Weeks
Google released Gemini 3.8 Flash, its third Flash model in six weeks, pricing it identically to its predecessor at $0.75 per million input tokens and $3.75 per million output tokens. The launch coincides with a favorable antitrust ruling and public praise from Berkshire Hathaway's Greg Abel, giving Alphabet a stronger narrative after a four-month stock slide.
Google's WeatherNext 3 Drops Physics Simulations, Learns Weather Forecasting Directly From Satellite Data
Google and DeepMind released WeatherNext 3, an AI weather model that trains directly on live geostationary satellite data instead of physics-based simulations. The model produces hourly forecasts at up to 5-kilometer resolution and now powers weather features in Google Search, Maps, and Gemini.
Comments
Loading...