arXiv論文メモ
新着一覧
cs.CL / cs.AI · 査読状況未確認

AIとの対話では人同士の協調の仕組みが逆転するか

Talking Past the Machine: Morality, Politeness, and Alignment in Human-AI Dialogue

Marina Mitiaeva, Lu Xiao

この論文をやさしく読む

ひとことで言うと

丁寧で温かいAIの応答が、人同士と同じような歩み寄りにつながるのかを対話データから調べた研究です。人間同士と人間・AIでは、同じ表現と同調の関係が逆になる場合が報告されています。

何に役立つ?

対話AIを一つの応答の印象だけで評価せず、会話全体の相互調整で評価する視点を与えます。ユーザーが会話を方向付ける余地を検討する手掛かりにもなります。

この研究の面白いところ

断定を避けたり表現を和らげたりすることが、人間同士では歩み寄りと関連する一方、AIでは同調の低下と関連しました。表面上の丁寧さと協調の成立を分けて検討しています。

どこまで分かった?

対話データと混合効果モデルに基づく関連の分析です。AIに社会的構造がないという解釈は著者らの結論であり、この分析だけで内的機構や表現の因果効果を確定したとは扱えません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

対話型AIは流暢で社会的に適切な応答を生成するが、協調的なコミュニケーションに参加しているのか、その表面的な形式を模倣しているだけなのかは明らかでない。これは、こうしたシステムの評価、信頼、設計に関わる中心的な問題である。本研究では、協調的対話の中心となる三つの側面、道徳性、丁寧さ、同調が、人間同士の会話と比べて人間とAIのやり取りでどのように働くかを調べる。人間とChatGPTの複数ターン対話15,881件と、人間同士の複数ターン対話10,784件を分析し、混合効果モデルによってターン間の同調を予測する特徴を特定する。一貫した乖離が観察された。AIは協調的なコミュニケーションの表面的特徴を生み出すが、その基礎にある社会的構造を伴っていない。道徳的な出力は交渉されるというより事前設定されているように見え、相手の面子への感受性を伴わずに温かみが生成され、言語的な収束は持続的に低下する。特に顕著なのは、協調の仕組み自体が逆方向に働くことである。人間同士ではより大きな歩み寄りに関連する断定を避ける表現や和らげる表現が、AIによる場合には同調の低下と関連する。また、人間同士では隔たりに関連する純潔性の枠付けが、ユーザーのAIへの収束と同時に現れる。ユーザーにやり取りを形作る余地を与える主体性は、両方の対話形式で同調の最も一貫した予測因子である。一方、より新しいモデルで道徳的な断定が弱くなっても、協調の改善は伴っていない。総合すると、これらのパターンは、AIが人間同士の協調の基盤となる相互適応を伴わずに協調の表面を再現していることを示唆する。さらに意外なことに、人間の歩み寄りを支える仕組みがAIでは逆向きに働き得る。このことは、ターン単位の見方だけでは、やり取り全体の成功を捉えるには不十分な可能性を示唆する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Conversational AI systems produce fluent, socially appropriate responses, yet whether they participate in cooperative communication or merely simulate its surface forms remains unclear - a question central to how these systems are evaluated, trusted, and designed. This study investigates how morality, politeness, and alignment - three dimensions central to cooperative dialogue - function in human-AI interaction compared to human-human conversation. We analyze 15,881 human-ChatGPT and 10,784 human-human multi-turn dialogues, using mixed-effects models to identify which features predict turn-to-turn alignment. We observe a consistent dissociation: AI produces the surface features of cooperative communication without the underlying social architecture. Moral output appears preconfigured rather than negotiated; warmth is generated without face sensitivity; linguistic convergence declines persistently. Most strikingly, the cooperative mechanisms themselves reverse direction: hedging and softening associated with greater accommodation between humans are associated with reduced alignment when produced by AI, and purity framing associated with human divergence coincides with users converging toward the AI. Agency - giving users room to shape the exchange - is the most consistent predictor of alignment across both interaction types, while lower moral assertiveness in more recent models is not accompanied by better cooperation. Together these patterns suggest that AI reproduces the surface of cooperation without the mutual adaptation that grounds it between humans - and, more surprisingly, that mechanisms sustaining human accommodation can run in reverse with AI, suggesting a turn-level view may be insufficient for interaction-level success.

著者のコメント

Accepted at the 60th Hawaii International Conference on System Sciences (HICSS-60)

arXiv ID: 2609.21401 / 要約の誤りについて