制御制約のあるHJI方程式の方策反復を解析する
Policy iteration for Hamilton-Jacobi-Isaacs equations with control constraints and comparison with Hamilton-Jacobi-Bellman equations
この論文をやさしく読む
ひとことで言うと
入力に制約がある制御問題で、HJI方程式を解く方策反復の収束を解析し、決定論的・確率的な場合を比較します。
何に役立つ?
制御制約を含む最適制御やゲーム理論的制御の数値計算を設計する際の理論的な参考になります。
この研究の面白いところ
一階と二階のHJIを扱い、数値試験ではHJBの一階・二階の解を、制約の有無で比較します。
どこまで分かった?
要旨には収束条件の詳細や数値比較の定量的結果は記載されていません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
制御入力に制約がある場合について、Hamilton-Jacobi-Isaacs(HJI)方程式を満たす価値関数を求める二段階の方策反復アルゴリズムの収束を解析する。決定論的な系に対応する一階のHJI方程式と、確率的な系に対応する二階のHJI方程式の両方を扱う。数値試験では、時間を逆向きに解くHJI偏微分方程式に半陰的な風上差分法を適用し、制御に制約がない場合とある場合の両方について、一階および二階のHamilton-Jacobi-Bellman(HJB)方程式の解を詳しく比較する。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Convergence of the bilevel policy iteration algorithm for the value function, satisfying the Hamilton-Jacobi-Isaacs (HJI) equation, is analyzed in the presence of control constraints. Both first and second-order HJI equations corresponding to the deterministic and stochastic systems are considered. For numerical tests, a semi-implicit upwind scheme for backward HJI PDEs is applied and a thorough comparison between the solutions to first and second-order Hamilton-Jacobi-Bellman (HJB) equations is presented for both the unconstrained and the constrained control cases.
著者のコメント
30 pages, 36 figures, Accepted in Computational and Applied Mathematics
arXiv ID: 2609.29368 / 要約の誤りについて