Black Forest Labs发布Flux 3:生成最长20秒带原生音频视频

Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs

精选理由

Black Forest Labs的Flux 3能生成带声音的视频,最长20秒,性能还略超Seedance 2.0,他们想搞世界模型,值得关注。

AI 摘要

Black Forest Labs发布Flux 3多模态基础模型,可同时学习图像、视频和音频。该模型能生成最长20秒带原生音频的视频,为该公司首次实现语音与画面同步输出。BFL内部测试显示Flux 3在多项指标上略领先Seedance 2.0,但独立评测尚未公开。公司最终目标构建世界模型,已在机器人任务上测试Flux 3。

原文 · Decoder

Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs

Black Forest Labs has released Flux 3, a multimodal foundation model that learns from images, video, and audio and can generate video with native sound for the first time. BFL's own tests put it just ahead of market leader Seedance 2.0, though independent results aren't yet available. The company ultimately wants to build a world model and is already testing Flux 3 on robotics tasks. The article Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs appeared first on The Decoder .

Black Forest Labs发布Flux 3:生成最长20秒带原生音频视频 · AI 热点