多模态推理开发者终于有了一个统一的高效引擎——vLLM-Omni 支持 30+ 模型和多种硬件,做多模态应用或推理优化的团队可以直接拿来用,省去重复造轮子的时间。
vLLM-Omni 项目在 GitHub 上达到 5000 星标,从去年 11 月社区启动至今,已发展为支持 30 多种多模态模型的高效推理引擎。它覆盖 Qwen3-Omni、HunyuanImage-3.0、Wan 2.2、BAGEL、MiMo-Audio 和 Flux2 等模型,并兼容 NVIDIA、AMD、华为昇腾、Intel 等多种硬件。该项目致力于提供可扩展、开源的多模态推理方案,吸引了大量社区贡献。
🚀 vLLM-Omni just hit 5K GitHub stars! 🎉 From a co…
🚀 vLLM-Omni just hit 5K GitHub stars! 🎉
From a community kickoff in November to powering omni-modality inference everywhere: vLLM-Omni supports 30+ models including Qwen3-Omni, HunyuanImage-3.0, Wan 2.2, BAGEL, MiMo-Audio, and Flux2, across NVIDIA, AMD, Huawei Ascend, Intel, and more.
Huge thanks to our amazing community, model partners, and everyone who's contributed a PR, filed an issue, or just given it a try.
Onto the next chapter of efficient, scalable, open multimodal inference.
⭐ https://t.co/JWmNKjhcRX 📚 https://t.co/xiwP0uR8UN
#vLLM #vLLMOmni