为视障人士设计的物体获取系统Touvigation在陌生室内环境中的表现
Touvigation: Embodied Adaptive Object Acquisition for Blind and Low-Vision Users in Unfamiliar Indoor Environments
这个系统很实用,专门为视障人士设计的,能帮助他们更快更准确地找到并拿到东西,比单纯用语言描述的助手效果更好。
Touvigation系统通过结合视觉语言理解和持续局部空间建模,为视障和低视力用户提供了低延迟、以身体为中心的引导。在12名参与者的测试中,该系统在物体获取任务上的成功率达到了100%,显著高于多模态大语言模型助手的58%和未辅助搜索的85%。
Touvigation: Embodied Adaptive Object Acquisition for Blind and Low-Vision Users in Unfamiliar Indoor Environments
Blind and low-vision users often face challenges when locating and physically acquiring objects in unfamiliar indoor environments. Existing vision-language-model-based assistants can provide semantic descriptions but may introduce latency, hallucinations, and guidance that is poorly aligned with embodied action. We present Touvigation, a hands-free object acquisition system that combines vision-language understanding with persistent local spatial modeling to provide low-latency, body-relative guidance. Drawing on formative interviews with eight blind and low-vision participants, we design a multi-stage guidance framework that adapts spatial references as users transition from orienting, to walking, to reaching and tactile verification. We evaluated Touvigation with 12 blind and low-vision participants against a multimodal large-language-model assistant and unassisted search. Touvigation achieved 100% task success, compared with 58% for the multimodal assistant and 85% for unassisted search, while reducing completion time and cognitive workload. Our findings demonstrate how persistent spatial grounding and adaptive embodied guidance can improve object acquisition for blind and low-vision users.