AI模型精选

谷歌发布 Gemini 3.5 Live Translate 实时语音翻译模型

Our latest audio model, Gemini 3.5 Live Translate, takes real-time speech translation to the next le...

精选理由

谷歌新模型,能实时翻译70+语言

AI 摘要

Gemini 3.5 Live Translate 是谷歌最新的音频模型,支持 70+ 语言的低延迟实时语音翻译。它通过流式处理语音,实现近实时的翻译输出,并具备多语言输入、自动语言检测、原生音频处理(保留语调、节奏和音高)以及噪声鲁棒性(在嘈杂环境中过滤背景噪音)等特点。开发者可利用该模型构建更自然的语音交互应用。

原文 · Google AI Developers

Our latest audio model, Gemini 3.5 Live Translate, takes real-time speech translation to the next le...

Our latest audio model, Gemini 3.5 Live Translate, takes real-time speech translation to the next level for developers by delivering low-latency translation across 70+ languages. By processing speech as it streams in near real time, the model enables devs to build low-latency audio experiences with: — Multilingual input: Understands multiple languages in a single session without needing to adjust settings. — Auto-detection: Identifies the spoken language and begins translation instantly. — Native audio processing: Generates more natural-sounding speech that preserves speakers' intonation, pacing, and pitch. — Noise robustness: Filters out ambient noise for clearer conversation in loud environments. Your browser does not support the video tag. 🔗 View on Twitter 💬 9 🔄 22 ❤️ 140 👀 21025 📊 30 ⚡