Meta 把自家的多模态模型 Muse Spark 用起来了,能生成字幕还能做内容分析,比之前版本更实用了。
Meta 内部部署多模态模型 Muse Spark。该模型与 Muse Image 协作进行智能体媒体生成,并为 Muse Image 和 Muse Video 生成详细字幕作为训练数据。Muse Spark 能将原始文本、图像和视频内容转化为下游应用可用的信号和洞察。Muse Spark 1.2 的评估和演示可在此查看:go.meta.me/multimodal。
Internally, Muse Spark is deployed in several areas including media generation. For example, the mod...
Internally, Muse Spark is deployed in several areas including media generation. For example, the model works with Muse Image for agentic media generation and produces detailed captions as training data for Muse Image and Muse Video. Muse Spark is also able to translate raw text, image, and video content into signals and insights that can be used in downstream applications. See more Muse Spark 1.2 evals and demos here: go.meta.me/multimodal Your browser does not support the video tag. 🔗 View on Twitter 💬 0 🔄 0 ❤️ 11 👀 1856 📊 2 ⚡