想做机器人深度估计?LingBot 团队开源了视觉基础模型和深度模型,误差直接减半,还拿下了12个基准的第一。
LingBot-Vision 作为视觉基础模型开源发布,其配套的 LingBot-Depth 2.0 在 150M 规模上训练,将深度估计误差降低一半,并在 12/16 个基准测试中达到领先水平。该模型专门解决了玻璃、镜面、透明物体等传统深度相机的难点。团队同时开源了这两项成果。
Cool open-source release. Physical AI is the next frontier, so things are shifting from thinking to...
Cool open-source release. Physical AI is the next frontier, so things are shifting from thinking to taking action. LingBot-Vision looks like a strong visual foundation model and shows progress in consistent depth on long, continuous operations. Robbyant @robbyant_brain 🪞 Glass. Mirrors. Transparent objects. — The nightmare of every depth camera. We just solved it! Introducing LingBot-Depth 2.0: 150M-scale training, half the depth error, 12/16 benchmarks topped. Powered by LingBot-Vision — the visual foundation model behind Depth's breakthrough. Both released today. LingBot-Vision is fully open-sourced. 🧵 #Robotics i #DepthEstimation i #OpenSource r #EmbodiedAI dAI Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 3 🔄 2 ❤️ 15 👀 3009 📊 5 ⚡