倾听与镜像:言语调谐与行为模仿对VR中具身AI代理的社会与共情感知的影响
Listening and Mirroring: The Effects of Verbal Attunement and Behavioral Mimicry on Social and Empathic Perceptions of Embodied AI Agents in VR
浏览论文内容
中文总结 AI 辅助
本研究通过VR中的具身AI咨询师实验,发现言语调谐是提升共情感知的关键,而行为模仿仅边缘影响类人性,表明多模态同步需谨慎设计。
中文摘要 AI 辅助
随着具身代理在VR中承担越来越多的社交和关系角色,仅靠视觉真实感和具身化可能不够;用户还必须感知这些代理在情感上协调、支持且类人。先前研究表明,言语调谐和非语言模仿各自都能改善用户对具身代理的社会评价。然而,行为模仿在很大程度上是在实时、对话式AI交互之外进行研究的,导致对当代理在沉浸式对话中同时生成上下文相关响应并调整其非语言行为时用户如何反应的理解有限。为弥补这一空白,我们开发了一个具身AI咨询师,它将对话式AI与实时面部表情和姿势模仿相结合,同时产生言语调谐或中性的回应。我们在一个2×2被试内研究中评估了该系统,涉及20名参与者,操纵了言语调谐和行为模仿。结果显示,言语调谐是感知共情最可靠的驱动因素。行为模仿与感知类人性呈边缘相关,而更大的模仿暴露显示出初步的探索性正相关,与共情、积极性和类人性相关,尤其是在女性参与者中。总之,这些发现表明,多模态同步并非设计VR中共情对话代理的简单加法策略,并强调了在实时交互中考虑言语和非言语行为如何结合的必要性。
英文摘要
As embodied agents take on increasingly social and relational roles in VR, visual realism and embodiment alone may be insufficient; users must also perceive these agents as emotionally attuned, supportive, and humanlike. Prior work suggests that verbal attunement and nonverbal mimicry can each improve users' social evaluations of embodied agents. However, behavioral mimicry has largely been studied outside of real-time, conversational AI interactions, leaving limited understanding of how users respond when an agent simultaneously generates contextually responsive dialogue and adapts its nonverbal behavior during an immersive conversation. To address this gap, we developed an embodied AI counselor that combines conversational AI with real-time facial-expression and posture mimicry, while producing either verbally attuned or neutral responses. We evaluated the system in a 2 X 2 within-subjects study with 20 participants, manipulating verbal attunement and behavioral mimicry. Results showed that verbal attunement was the most reliable driver of perceived empathy. Behavioral mimicry showed a marginal relationship with perceived humanness, while greater mimicry exposure showed preliminary, exploratory positive associations with empathy, positivity, and humanness, particularly among female participants. Together, these findings show that multimodal synchrony is not a simple additive strategy for designing empathic conversational agents in VR and underscore the need to consider how verbal and nonverbal behaviors are combined during real-time interaction.
发表机构
- Drexel University(德雷塞尔大学)
机构由 AI 辅助整理,请以论文原文为准。