Meta把图像生成做成了agent,能调用工具和自我改进,还能和Muse Spark配合做多媒体,挺有意思的。
Meta推出Muse Image和Muse Video两款媒体生成模型,由Meta Superintelligence Labs开发。Muse Image采用agent工作流,能调用工具、自我精炼,并随测试时计算扩展而改进。它与Muse Spark协作生成多媒体内容。Muse Video基于同一预训练基础,支持原生音频。用户可在Meta AI应用、网页、Instagram Stories和WhatsApp中体验Muse Image,目前限部分国家。
Muse Image works as an agent rather than a direct prompt-to-image model: it invokes tools, self-refi...
Muse Image works as an agent rather than a direct prompt-to-image model: it invokes tools, self-refines, improves with scaled test-time compute, and pairs with Muse Spark for collaborative media generation. 🧵👇 AI at Meta @AIatMeta Introducing Muse Image and Muse Video, the first media generation models developed by Meta Superintelligence Labs. Muse Image is our most advanced image generation model yet. It follows instructions faithfully, edits with precision, composes from multiple references, and draws on Instagram for social context. It also brings agentic tool use capabilities to image generation and integrates with Muse Spark. You can try Muse Image in the Meta AI app and web, as well as in Instagram Stories and WhatsApp – starting in limited countries with more locations on the way. Today we’re also previewing Muse Video, which is built upon the same pretraining base as Muse Image to deliver exceptional visual fidelity with native audio support. Learn more about both models: go.meta.me/7f4427 Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 7 🔄 17 ❤️ 191 👀 25340 📊 28 ⚡