不完全な均衡データから効用を学ぶための損失関数
Suboptimality Loss for Inverse Learning from Imperfect Equilibria
この論文をやさしく読む
ひとことで言うと
複数主体の行動が完全な均衡でなくても、背後にある効用をデータから推定する。
何に役立つ?
競争や戦略的な行動の予測、反事実分析、制度設計で、雑音に強い推定法を考える参考になる。
この研究の面白いところ
一方的に行動を変えた場合の利得を損失とし、凸性と分解可能性を示す。
どこまで分かった?
実証例はネットワーク型Cournot競争であり、あらゆるゲームで同じ精度を保証するものではない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
現代の多くのシステムでは複数の主体が戦略的に影響し合う。観測された行動は、部分的にしか分からない効用の下での均衡行動を反映することが多い。データから隠れた効用を復元する逆ゲーム理論は、予測、反事実の分析、仕組みの設計に重要である。しかし、逆変分不等式に基づく従来の方法は、雑音を含む観測や、互いに整合しない均衡の観測に非常に敏感であり、利用範囲が制限される。 この問題に対し、観測された戦略の組から各プレイヤーが一方的に戦略を変えたときに得られる効用の増加を合計する、ゲーム理論的な劣最適性損失を導入する。第一に、この損失は凸であり、プレイヤーごとの最善応答へ効率的に分解できることを示す。第二に、予測可能性の損失と逆変分不等式の損失の間に挟まれることを示し、均衡予測の扱いやすい代替目標となる。第三に、これを最小化するミラー降下法を開発する。異質な主体からなるネットワーク型Cournot競争では、雑音のある観測と不整合な均衡データの下でも提案法は正確さを保ち、逆変分不等式の方法は退化した推定値を出すことを示した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-19(UTC)
- 最新改訂
- 2026-09-19 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-19 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Many modern systems involve the strategic interaction of multiple agents. In such settings, observed actions typically reflect equilibrium behavior under utilities that are only partially known. Recovering these hidden utilities from data - the central goal of inverse game theory - is key for prediction, counterfactual analysis, and mechanism design. However, existing approaches based on inverse variational inequalities are highly sensitive to noisy and inconsistent equilibrium observations, thus limiting their applicability. In this paper, we resolve this issue by introducing a game-theoretic suboptimality loss that measures the aggregate utility gain players could obtain by unilaterally deviating from an observed strategy profile. First, we show that this loss is convex and admits an efficient decomposition into player-wise best-responses. Second, we show this loss is sandwiched between the predictability loss and the inverse variational inequality loss, making it a tractable surrogate for equilibrium prediction. Third, we develop a mirror descent algorithm to minimize it and demonstrate on a heterogeneous networked Cournot competition that our approach remains accurate under noisy observations and inconsistent equilibrium data while inverse variational inequality methods produce degenerate estimates.
arXiv ID: 2609.23200 / 要約の誤りについて