Google Releases Gemini 3.5 Live Translate
Google has released Gemini 3.5 Live Translate, an audio model for live speech-to-speech translation across more than 70 languages.
The model is being positioned as infrastructure for real-time conversation rather than a batch translation tool. Google says it can begin translating as a person speaks, continue listening while producing output, and stay only seconds behind the speaker. The company also says the system can preserve pacing, pitch, and intonation over longer sessions, which is a harder target than producing a literal transcript in another language.
The release spans several Google surfaces. The company says Gemini 3.5 Live Translate is coming to Google AI Studio, Google Translate, and Google Meet, while the Google AI launch post specifically points users to the Google Translate app on iOS and Android.
The practical significance depends on latency and reliability in real conversations. Speech translation systems often look strong in controlled demos but struggle when speakers interrupt each other, use regional accents, mix languages, or move through noisy rooms. Google is making a broader claim here: that an audio model can handle simultaneous listening and speaking well enough for fluid conversation.
For developers, the Google DeepMind Gemini Audio page frames 3.5 Live Translate as the Gemini Audio option best suited to real-time speech-to-speech translation. That makes this release part of the wider shift toward multimodal models that handle voice interaction directly, instead of routing speech through separate transcription and text-translation stages.