arXiv論文メモ
新着一覧
math.OC · 査読状況未確認

不確かな流れの中で移動と観測を同時に最適化

The Price of Covertness: Dual-Control Navigation in Uncertain Flows under Adversarial Sensing

Ruimeng Hu, Botao Jin, Xu Yang

この論文をやさしく読む

ひとことで言うと

流れを学ぶための移動と、目的地へ進むための移動が互いに影響する状況を、検知されやすさも含めて数理モデルにした研究です。

何に役立つ?

移動計画と情報獲得を別々に扱うと見落とすトレードオフや、情報を得る順番の重要性を分析する理論になります。

この研究の面白いところ

同じ量の情報でも、早く得れば後の補正動作を減らせますが、後で得ても既に発生した漏洩は取り消せないという非対称性を感度公式で表します。

どこまで分かった?

小誤差展開や定式化したゲームの条件に基づく理論と数値実験です。実環境での移動実験や、あらゆる流れ・観測方式での秘匿性の保証は要旨にはありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

不確かな流れの中での秘匿移動では、経路計画と情報獲得が結び付いている。移動体は進むために局所的な流れを推定する必要があるが、推定誤差は補正操作を引き起こし、それが統計的な識別しやすさ、すなわち漏洩を増やして、移動体の検知可能性を高める。本研究では、位置に推定誤差の分散を追加することで、この結び付きをモデル化する。分散の変化は、局所的な情報獲得率を通じて経路に依存する。小誤差の展開により、推定の不確実性が期待検知率に主要次数の補正を与えることを示す。 得られる経路計画問題は、1階のハミルトン–ヤコビ–ベルマン(HJB)方程式で特徴付けられる。期限を添字とする価値関数が、秘匿性と時間のフロンティアを決める。また、センシング品質に関する厳密な感度公式から、経路の早い段階で得た情報は後に発生する漏洩を減らせるが、後から得た情報で既に生じた漏洩を減らすことはできないと分かる。 次に、センシングの配分をめぐるゼロ和ゲームを定式化し、漏洩率が配分について凸であることを示す。そして、防御側は単一地点に全配分する極点の間のランダム化に対象を限定してよいことを証明する。移動体側の最適応答をHJB方程式から求め、列生成法で均衡を計算する。数値実験により、学習のための迂回、期限と漏洩のトレードオフ、空間的な順序の効果、センシング配分のランダム化の利点を例示する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Covert navigation in an uncertain flow couples motion planning with information acquisition. A vehicle must estimate the local flow to navigate, while estimation error induces corrective maneuvers that increase statistical distinguishability, or leakage, and hence the vehicle's detectability. We model this coupling by augmenting position with the estimation-error variance, whose evolution depends on the route through the local information rate. A small-error expansion shows that estimation uncertainty contributes a leading-order correction to the expected detectability rate. The resulting route-planning problem is characterized by a first-order Hamilton--Jacobi--Bellman (HJB) equation. The deadline-indexed value determines the covert-time frontier, while an exact sensitivity formula with respect to sensing quality shows that information acquired earlier along the route can reduce leakage incurred later, whereas information acquired later cannot reduce leakage already incurred. We then formulate a zero-sum sensing-allocation game, show that the leakage rate is convex in the sensing allocation, and prove that the defender may restrict attention to randomization over extreme single-site allocations. The resulting equilibrium is computed by column generation, with vehicle best responses obtained from the HJB equation. Numerical experiments illustrate the learning detour, the deadline--leakage tradeoff, the spatial-ordering effect, and the benefit of randomized sensing allocations.

著者のコメント

8 pages, 3 figures

arXiv ID: 2609.24225 / 要約の誤りについて