他者への配慮の違いを与えるとAIのゲーム行動は人に近づくか
Conditioning LLMs on Social Value Orientation improves behavioural alignment in a sequential social dilemma
この論文をやさしく読む
ひとことで言うと
自分と相手の利益をどれだけ重視するかという人の違いをLLMに与え、ゲーム中の協力行動が人の分布に近づくかを調べています。8モデルで、条件付けが行動を体系的に変えました。
何に役立つ?
人の意思決定をLLMで模擬するとき、単一の平均的な人物像ではなく、選好の違いを入力する方法の評価に役立ちます。検証されたのはセンチピードゲームの2変種です。
この研究の面白いところ
人に近い行動を出せることと、モデル自身に安定した社会的選好があることを分けています。選好を行動へ変換できる能力が、模擬の有用性につながるという解釈です。
どこまで分かった?
整合性の改善は最大70%で、要旨にはその評価指標の具体的な定義はありません。プロンプトや提示順序への感度も報告されており、他の社会状況や安定した人格の再現まで実証したものではありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
大規模言語モデル(LLM)は人の意思決定のシミュレーションに使われることが増えているが、その出力はしばしば人間の行動の多様性を十分に表さない。本研究では、社会的価値志向性(SVO:自分の結果を他者の結果に対してどの程度重視するかを表す尺度)でLLMを条件付けることで、逐次的な社会的ジレンマにおける人の行動をよりよく再現できるかを調べる。 センチピードゲーム(CG)の2つの変種で得た実験データを参照し、8つのLLMの既定の行動と、人の標本から得たSVOプロファイルで条件付けした後の行動を比較する。既定の戦略はモデル間で大きく異なるが、概して人の参照分布から離れていることが分かった。一方、SVOプロファイルによる条件付けは戦略的行動を体系的に変化させ、人の参照データとの整合性を既定の場合と比べて最大70%改善した。モデルを通じて、誘導するSVO値を高くするとゲームを停止する確率が低下し、人の行動で観測される向社会性と協力の関係が再現された。 LLMから引き出されるSVOがプロンプトや順序の影響を受けることも合わせると、これらの結果は、行動シミュレーションにおけるSVOの有用性が、LLMが安定した社会的選好を持つことではなく、社会的選好を対応する戦略的選択へと写し取る能力に依存することを示唆する。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Large language Models (LLMs) are increasingly used to simulate human decision-making, yet their outputs often under-represent human behavioural heterogeneity. We investigate whether conditioning LLMs on Social Value Orientation (SVO, a measure of how individuals value their own outcomes relative to others') can better reproduce human behaviour in a sequential social dilemma. Using experimental data from two variants of the Centipede Game (CG) as reference, we compare the default behaviour of eight LLMs with behaviour generated after conditioning them on SVO profiles drawn from the human sample. We find that the default strategies vary substantially across models, but are generally distant from the reference distributions. However, conditioning them on SVO profiles systematically steers their strategic behaviour, improving alignment with the human reference by up to 70% with respect to the default. Across models, higher induced SVO values decrease the probability of stopping the game, reproducing the relationship between prosociality and cooperation observed in human behaviour. Together with the sensitivity of LLMs' elicited SVO to prompt and order effects, these results suggest that the usefulness of SVO for behavioural simulation does not depend on LLMs possessing stable social preferences, but rather on their ability to map social preferences onto corresponding strategic choices.
arXiv ID: 2610.01667 / 要約の誤りについて