跨语言沟通的痛点终于被解决了——Gemini 3.5 Live Translate 让实时翻译不再有尴尬停顿,经常需要与外语人士交流的团队或个人可以直接在 Google Translate 应用中体验。
Google AI 发布了 Gemini 3.5 Live Translate,这是其最新的音频模型,专为实时语音到语音翻译设计。该模型支持超过 70 种语言,能在用户开始说话的同时进行翻译,并流式输出结果,无需等待或停顿。它通过同时接收输入和输出翻译语音,在速度和翻译质量之间做出毫秒级决策,保持对话的流畅自然。此外,模型还能在长时间会话中维持语速、音高和语调,提升用户体验。目前该功能已在 Google Translate 应用的 iOS 和 Android 版本中上线。
Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech trans...
Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages and starts translating as soon as you start talking, streaming translations while listening to what you say next. No awkward pauses or choppy audio, just real connection without language barriers. So, how does it work? 🤔 The model is able to make split-second decisions to juggle speed and translation quality so conversations actually feel fluid, human, and natural. In order to do this, the model must receive and contextualize the input while simultaneously outputting the translated speech. Through this process, Gemini 3.5 Live Translate manages to stay mere seconds behind each speaker and can even maintain pacing, pitch, and intonation across extended sessions. See it in action below, or try it yourself in the Google Translate app on iOS & Android. Your browser does not support the video tag. 🔗 View on Twitter 💬 53 🔄 160 ❤️ 1190 👀 67713 📊 205 ⚡