VRの対話型AIで言葉の共感と動作の模倣を比較
Listening and Mirroring: The Effects of Verbal Attunement and Behavioral Mimicry on Social and Empathic Perceptions of Embodied AI Agents in VR
この論文をやさしく読む
ひとことで言うと
VRのAI相談役について、共感的な言葉と表情・姿勢の模倣が利用者の印象にどう関わるかを比べた研究。
何に役立つ?
考えられる用途は、VRで会話するエージェントの対話設計と評価。要旨で安定して示されたのは、共感的な言葉の効果である。
この研究の面白いところ
20人が四つの組み合わせを体験する設計で、言葉と動作の効果を分けて調べた。動作模倣の効果は言葉の効果ほど明確ではなかった。
どこまで分かった?
動作模倣と人間らしさの関係は境界的で、接触量との関連は探索的な結果。参加者20人の研究であり、幅広い利用者への効果は要旨だけでは確定できない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
VRで身体を持つエージェントが社会的・対人的な役割を担うようになると、見た目の写実性や身体表現だけでは足りず、利用者が感情を理解してくれる、支えになる、人間らしいと感じることも重要になる。先行研究は、言葉による歩調合わせと非言語的な模倣がそれぞれ社会的評価を高めうると示している。しかし、動作の模倣は主にリアルタイムの対話型AI以外で研究されてきたため、状況に応じた会話と非言語行動の適応を没入型の会話で同時に行った際の反応は十分に分かっていない。 この課題に対し、対話型AIとリアルタイムの表情・姿勢の模倣を組み合わせ、言葉では共感的な応答または中立的な応答を返す身体を持つAIカウンセラーを開発した。参加者20人がすべての条件を体験する2×2の被験者内研究で、言葉による歩調合わせと動作模倣を操作して評価した。 知覚された共感を最も安定して高めたのは、言葉による歩調合わせだった。動作模倣と人間らしさの評価の関係は境界的で、模倣への接触が多いほど共感、好意的な評価、人間らしさが高いという関連は予備的・探索的に見られ、特に女性参加者で目立った。これらの結果は、複数の表現を同期させる効果が単純に足し合わせられるわけではなく、リアルタイムの対話では言語行動と非言語行動の組み合わせ方を考える必要があることを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-23(UTC)
- 最新改訂
- 2026-09-23 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-23 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
As embodied agents take on increasingly social and relational roles in VR, visual realism and embodiment alone may be insufficient; users must also perceive these agents as emotionally attuned, supportive, and humanlike. Prior work suggests that verbal attunement and nonverbal mimicry can each improve users' social evaluations of embodied agents. However, behavioral mimicry has largely been studied outside of real-time, conversational AI interactions, leaving limited understanding of how users respond when an agent simultaneously generates contextually responsive dialogue and adapts its nonverbal behavior during an immersive conversation. To address this gap, we developed an embodied AI counselor that combines conversational AI with real-time facial-expression and posture mimicry, while producing either verbally attuned or neutral responses. We evaluated the system in a 2 X 2 within-subjects study with 20 participants, manipulating verbal attunement and behavioral mimicry. Results showed that verbal attunement was the most reliable driver of perceived empathy. Behavioral mimicry showed a marginal relationship with perceived humanness, while greater mimicry exposure showed preliminary, exploratory positive associations with empathy, positivity, and humanness, particularly among female participants. Together, these findings show that multimodal synchrony is not a simple additive strategy for designing empathic conversational agents in VR and underscore the need to consider how verbal and nonverbal behaviors are combined during real-time interaction.
著者のコメント
11 pages. Accepted to ACM VRST 2026
arXiv ID: 2609.27246 / 要約の誤りについて