Black Forest Labs 的 FLUX 3 能用一个提示生成20秒多镜头真实视频,还可以做图像、音频和机器人预测,很值得试试。
Black Forest Labs 推出 FLUX 3,一个统一多模态模型,支持图像、视频、音频和动作预测。其视频生成能力突出,单次提示即可生成最长20秒的多镜头视频,细节逼真。目前 FLUX 3 Video 已开放早期访问。该模型采用统一架构训练,可扩展至机器人动作预测。
I've been playing around with FLUX 3 for the last few week - it's an incredibly impressive model. I...
I've been playing around with FLUX 3 for the last few week - it's an incredibly impressive model. It can generate up to 20 seconds of video across multiple shots with one prompt 🤯 This model is so strong at the little details that make a clip look real - a few favorites 👇 Your browser does not support the video tag. 🔗 View on Twitter Black Forest Labs @bfl_ai Introducing FLUX 3. One multi-modal model for Image, Video, Audio and Action-Prediction. Creations are truer to life in every kind of style. FLUX 3 Video is now available in early access (link below). Jointly trained in one unified architecture, our model can be extended to predict actions for robotics. See our work with mimic and Audi in the thread. Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 5 🔄 1 ❤️ 12 👀 1210 📊 5 ⚡
- 歸藏(guizang.ai)15:43原文