实时语音翻译终于不再是“等说完再翻”的延迟体验——做跨国会议、直播或外语学习的人可以直接用上,建议试试 Gemini Live API 或 Google Translate 的更新。
Google 发布了 Gemini 3.5 Live Translate,一种实时语音到语音翻译模型。与等待完整句子的传统系统不同,它能在说话人仍在讲话时就开始翻译,通过流式翻译技术预测并更新翻译内容。该模型支持 70 多种语言,延迟仅几秒,并能保留语速、音调和语调。它已通过 Gemini Live API、Google Meet 预览版以及 Android/iOS 上的 Google Translate 向用户推出。
Fascinating. Google just released Gemini 3.5 Live …
Fascinating. Google just released Gemini 3.5 Live Translate.
A live speech-to-speech translation model that starts speaking in another language while the original speaker is still talking.
Older translation systems often wait for a full sentence, because early words can be misleading until later words reveal tense, intent, or context.
Gemini 3.5 instead runs streaming translation, where the model listens, interprets partial meaning, predicts what can safely be translated, and keeps updating as new speech arrives.
supports 70+ languages, stays only a few seconds behind the speaker, and can preserve pacing, pitch, and intonation across longer sessions.
Rolling out to Gemini Live API, businesses through Google Meet preview, and regular users through Google Translate on Android and iOS.