中国AI实验室SenseNova U1发布:原生多模态建模,38B激活3B MoE

Chinese AI labs are increasingly releasing very se…

精选理由

SenseNova U1解决了多模态生成中模块切换导致的信息丢失问题,做信息图、海报、漫画等密集视觉内容的创作者可以直接用ComfyUI体验,效果惊艳。

AI 摘要

中国AI实验室商汤科技在HuggingFace上发布了SenseNova U1模型,采用原生多模态建模和MoT架构(38B激活3B MoE)。该模型将多模态生成视为一个统一的建模问题,而非分离的视觉、语言和图像模块链,从而减少了模块间的信息损失,提升了生成内容的一致性。SenseNova U1特别擅长生成可读、结构化、一致的图文输出,如信息图、指南、海报、漫画等。它支持ComfyUI,推理速度快(A3B),为密集视觉内容创作提供了高效工具。

原文 · rohanpaul_ai

Chinese AI labs are increasingly releasing very se…

Chinese AI labs are increasingly releasing very serious open source work.

SenseNova U1 just dropped on HuggingFace: native multimodal modeling, MoT architecture (38B-Active 3B MoE)

It attacks the hardest part of image generation: readable, structured, consistent image-text output.

The most interesting part of SenseNova U1 is it treats multimodal generation as one native modeling problem, not a chain of separate vision, language, and image modules.

That means less handoff between modules, less information loss, and better consistency when creating dense visual content like infographics, guides, posters, comics, and image-text workflows.

ComfyUI support, fast A3B inference, and absolutely brilliant for dense visuals like infographics, posters, comics, and guides.