arXiv論文メモ
新着一覧
eess.SY / cs.SY · 査読状況未確認

単位時間当たりの報酬を最適化する混雑ゲーム

Reward-Rate Congestion Games and Replicator--Dinkelbach Dynamics

Hassan Abdelraouf, Vaibhav Srivastava, and Vijay Gupta

この論文をやさしく読む

ひとことで言うと

複数の主体が混雑を考慮しながら、単位時間当たりの報酬を最大にするゲームと更新則を考えた理論研究。

何に役立つ?

ロボットやサイバー・フィジカルシステムのタスク配分で、個々と全体の報酬率をどう最適化するか考える枠組みになる。

この研究の面白いところ

元のゲームをDinkelbach変換でポテンシャルゲームにし、有限終了や更新則の安定性を条件付きで示した。

どこまで分かった?

有限終了には内側の最大化を大域的に解く条件があり、連成系の局所指数安定性には十分に遅い更新が必要。要旨の例は連続タスク配分である。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

報酬率は、時間、作業負荷、協調のコストが限られた資源となるサイバー・フィジカルシステムやロボットシステムで重要な性能基準である。本研究は、各エージェントが実行時間当たりの報酬を最大化しようとする報酬率混雑ゲームを導入する。直接の報酬率ゲームは一般に厳密ポテンシャルゲームではない。そこでDinkelbach法に基づく枠組みを作り、Dinkelbachパラメータを固定した変換後のゲームが厳密ポテンシャルゲームとなるようにする。内側のポテンシャル最大化問題を大域的に解く場合、ポテンシャル水準のDinkelbach反復は、最適なポテンシャル報酬率で有限回のうちに終了する。また、変換後のゲームの均衡が、元の報酬率ゲームの均衡でもあるための十分条件を示す。 集団全体の性能を最適化するため、限界的な外部性を補正し、補正後のポテンシャルがDinkelbach変換した社会的報酬率の目的関数と一致するようにする。これによって社会的報酬率を最適化できる。さらに、報酬率に基づく集団ゲームについて、速い複製子ダイナミクスと遅い報酬率更新を結び付けた連続時間の複製子・Dinkelbachダイナミクスを開発する。パラメータ固定時の複製子ダイナミクスの収束、縮約したDinkelbachダイナミクスの大域的な漸近安定性と局所的な指数安定性、ならびにDinkelbach更新が十分に遅いときの連成系の局所的な指数安定性を確立する。この枠組みを連続的なタスク配分問題で例示する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Reward rate is a key performance criterion in cyber-physical and robotic systems where time, workload, and coordination costs are limiting resources. We introduce reward-rate congestion games, where agents seek to maximize reward per unit execution time. The direct reward-rate game is generally not an exact potential game. We develop a Dinkelbach-based framework in which, for every fixed Dinkelbach parameter, the transformed game is an exact potential game. This yields a potential-level Dinkelbach iteration that terminates finitely at the optimal potential reward rate when the inner potential maximization problem is solved globally. We also provide a sufficient condition under which an equilibrium of the transformed game is an equilibrium of the original reward-rate game. To optimize aggregate performance, we introduce marginal externality corrections that make the corrected potential coincide with the Dinkelbach-transformed social reward-rate objective, thereby enabling optimization of the social reward rate. Finally, we develop a continuous-time replicator--Dinkelbach dynamics for reward-rate population games coupling fast replicator dynamics with a slow reward-rate update. We establish convergence of the fixed-parameter replicator dynamics, global asymptotic and local exponential stability of the reduced Dinkelbach dynamics, and local exponential stability of the coupled system for sufficiently slow Dinkelbach updates. The framework is illustrated on a continuous task-allocation problem.

arXiv ID: 2609.28240 / 要約の誤りについて