OpenAI 出了两个新转录模型,一个实时低延迟,一个异步批量处理,对口音和噪音场景更准。做语音识别应用的话可以试试。
OpenAI 在 API 中推出两个新转录模型:GPT-Live-Transcribe 专为低延迟实时转录设计,GPT-Transcribe 优化了已完成音频文件的异步转录和批量处理。这两个模型能更好理解上下文,在口音、多种语言、短句、数字、专业术语以及嘈杂背景下的语音识别准确率更高。
We're introducing two new transcription models in the API: • GPT-Live-Transcribe: built for low-lat...
We're introducing two new transcription models in the API: • GPT-Live-Transcribe: built for low-latency live transcription. • GPT-Transcribe: optimized for asynchronous transcription of completed audio files and batch workloads. Both models better understand context and deliver more accurate transcription on real world audio across accents and languages, including for short phrases, numbers, specialized terminology, and speech with loud background noise. Your browser does not support the video tag. 🔗 View on Twitter 💬 20 🔄 23 ❤️ 144 👀 7765 📊 44 ⚡