论文

VoCa:让语音智能体在对话中配合画布协同表达

VoCa: Designing Speech-Canvas Interaction for Voice-Based Conversational Agents

精选理由

一个会一边说话一边画图的语音智能体 VoCa,18 个人实测五天,想知道语音交互还能怎么玩可以看看这篇。

论文 VoCa 研究语音智能体如何在多轮对话中同时使用语音和画布。研究先通过一项观察实验分析两人如何协调口头讲解与板书,再通过设计工作坊确定语音-画布交互的设计空间。在此基础上开发的 VoCa 能将语音与视觉对象的创建、标注和注意力引导结合起来。为期五天、18 名参与者的部署测试总结了学习、工作和日常场景中的使用模式,以及协调智能体说什么与展示什么所面临的挑战。

原文 · arXiv cs.AI

VoCa: Designing Speech-Canvas Interaction for Voice-Based Conversational Agents

People write and sketch while speaking to explain, organize, and develop content together. Inspired by these practices, we investigate how voice agents can use a canvas alongside speech in multi-turn conversations with users. We conducted a two-part formative study: an observational study of how pairs coordinated speech and boardwork, followed by a design workshop that informed a design space for speech-canvas interaction with voice agents. Building on these insights, we developed VoCa, a voice agent that coordinates speech with visual object creation, annotation, and attention guidance. A five-day deployment with 18 participants examined usability, experiences of speech-canvas interaction, patterns of use, and desired improvements. Participants' experiences highlighted opportunities for speech-canvas interaction in learning, work, and daily life, alongside challenges in coordinating what agents say and show in ways users can follow and influence. These findings inform how voice agents can use a canvas alongside speech in conversation.