Gemini 3.8 TTS 支持录音 20 秒复刻声音,或用提示词设计全新音色
Philipp Schmid 演示了 Gemini 3.8 TTS 怎么用 20 秒录音复刻你的声音,或者用提示词造一个全新音色,流程就三步。
Philipp Schmid 介绍了 Gemini 3.8 TTS 的语音复刻流程:录制 20 秒语音加一句同意声明,通过 API 调用创建自己的声音。创建后的声音可在任意请求中使用,语气风格通过 speech_metadata 字段控制。除了复刻本人声音,也可以直接用提示词设计一个完全自定义的音色。文章还提供了让智能体自动完成录音、创建声音和生成测试语音的提示词。
With Gemini 3.8 TTS you can replicate your voice or design a completely custom one from a prompt 1. Record 20s of you talking + the consent sentence 2. Create your Voice via API call 3. Use it in any request, style goes in speech_metadata Past this into your agent "Read philschmid.de/gemini-3-8-tts and walk me through creating my own voice for Gemini 3.8 TTS. Check my setup first (GEMINI_API_KEY, ffmpeg, gemini-skills), help me record the two clips, create the voice, and generate a test line I can listen to." or read below. Philipp Schmid @_philschmid x.com/i/article/2103… 🔗 View Quoted Tweet 💬 3 🔄 1 ❤️ 19 👀 1745 📊 5 ⚡