標本経路だけを使うランダムウォーク停止問題の最適保証
Sample-Based Prophet Inequalities for Random Walks
この論文をやさしく読む
ひとことで言うと
値がランダムに変わる過程で、過去に見た有限本の経路だけを使い、どの時点で止めれば全経路を知る理想的な選択に近づけるかを調べます。
何に役立つ?
未知の分布の下で逐次的な停止判断を行う際、標本数に応じてどこまでの期待報酬を保証できるかを理解する基礎理論です。
この研究の面白いところ
標本数Kと保証の関係を明示的な式で示し、標本が増えると完全情報の場合の1/eに近づくことを示しています。
どこまで分かった?
増分が独立同分布のランダムウォークという設定です。定数は期待値の比の保証で、個々の経路で最大値の同じ割合を必ず得られるという意味ではありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
標本に基づく情報を使う、ランダムウォーク報酬の停止問題について、予言者不等式を研究する。目的は、独立同分布の増分を持つランダムウォークで、最大値にできるだけ近いところで停止することである。性能は、停止時の期待報酬と、真の最大値の期待値との比で測る。増分の分布は未知で、意思決定者が報酬過程の独立なK本の標本経路を利用できるモデルを考える。無限期間の設定では、定数(K/(K+1))^(K+1)を持つ最良の予言者不等式を確立する。この保証は、ランダムウォークのラダー高分解に基づく無作為化停止規則によって達成される。K → ∞では、完全情報の設定における古典的な1/eの予言者不等式が再現される。nステップ後に過程が終了する有限期間の設定では、まず情報なしの場合について、定数1/Hₙを持つ最良の予言者不等式を証明する。ここでHₙは第n調和数である。K ≥ 1本の標本がある場合、予言者定数1/4が達成可能であることを示す。最後に、K本の標本では、n ≥ 2K²のとき予言者定数が高々(K/(K+1))^(K+1) + (6+6Hₖ)/Hₙであることを証明する。これは、n → ∞で無限期間の定数へ収束することを意味する。この方法は、ラダー高やSpitzerの恒等式を含むランダムウォーク理論と、線形計画の双対性を組み合わせる。本結果は、ランダムウォークの停止理論と、まだほとんど研究されていない相関した報酬に対する標本ベースの予言者不等式に貢献する。著者らの知る限り、今回の最良の標本ベース予言者不等式は、利用できる標本数によって性能を定数倍の精度だけでなく厳密にパラメータ化した初めてのものである。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-22(UTC)
- 最新改訂
- 2026-09-22 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-22 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We study prophet inequalities for a random walk reward stopping problem with sample-based information. The goal is to stop as close as possible to the maximum of a random walk with i.i.d. increments, measuring performance by the ratio between the expected reward when stopping and the expected true maximum. We consider a sample-based model in which the increment distribution is unknown and the decision maker has access to $K$ independent sample paths of the reward process. For the infinite-horizon setting, we establish a sharp prophet inequality with constant $(K/(K+1))^{K+1}$. The guarantee is attained by a randomised stopping rule based on the ladder height decomposition of random walks. As $K\to\infty$, this recovers the classical $1/e$ prophet inequality from the full-information setting. For the finite-horizon setting, where the process terminates after $n$ steps, we first prove a tight no-information prophet inequality with constant $1/H_n$, where $H_n$ is the $n$-th harmonic number. For $K\ge1$ samples, we show that a prophet constant of $1/4$ is attainable. Finally, we prove that, with $K$ samples, the prophet constant is at most $(K/(K+1))^{K+1}+(6+6H_K)/H_n$ for $n\ge 2K^2$, implying convergence to the infinite-horizon constant as $n\to\infty$. Our approach combines random walk theory, including ladder heights and Spitzer's identity, with linear programming duality. Our results contribute to random walk stopping theory and to sample-based prophet inequalities for correlated rewards, an area that remains largely unexplored. To the best of our knowledge, our tight sample-based prophet inequalities are the first whose performance is parameterised exactly, rather than only up to constants, by the number of available samples.
著者のコメント
Accepted at the 22nd Conference on Web and Internet Economics (WINE 2026)
arXiv ID: 2609.26017 / 要約の誤りについて