arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

人混みを走るロボットに障害物の不確かさと回復行動を追加

DUGM-R: Uncertainty-Aware Dynamic Grid Mapping and Risk-Triggered Recovery for Learned Local Navigation

Haoyun Feng, Adrian Rubio-Solis, Zhaodong Guo, George Mylonas

この論文をやさしく読む

ひとことで言うと

人や物が動く屋内で、障害物の動きの不確かさを地図に入れ、危険時に回復行動へ切り替える方法です。

何に役立つ?

混雑した屋内を走る移動ロボットの衝突リスクを下げる設計に役立つ可能性がある。

この研究の面白いところ

通常方策を固定した後で、衝突予測と回復用の方策を別に追加する。

どこまで分かった?

要旨の評価は臨床物流のシミュレーションとTurtleBot3への導入であり、改善幅の数値や長期実運用の結果はない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

混雑した屋内で学習によって局所移動するロボットの性能は、動く障害物をどう表現するかに左右される。また、通常の方策を学習した後も衝突しやすい行動が残ることがある。著者らは、不確かさを考慮する動的格子地図DUGMと、学習後に追加する独立した回復機構によって、この二つの問題に対処するリスク考慮型の強化学習枠組みを提示する。DUGMは、局所的な占有状態、推定した障害物の動き、その推定の不確かさを、ロボット中心の表現にまとめる。通常方策を固定した後、その方策での走行結果から、有限期間のリスク価値関数RVFを学ぶ。通常方策を続けると衝突しやすいと予測された場合には、専用の回復方策を起動する。学習には使わなかったNVIDIA Isaac Simの臨床物流ベンチマークでは、不確かさを含む動的表現が静的または決定論的な代替表現より通常の移動を改善し、回復機構が残る衝突しやすい行動をさらに減らした。全体の枠組みは、方策の微調整、再学習、場所ごとの適応をせずにTurtleBot3にも直接導入され、シミュレーションで見られた性能の傾向を保った。結果は、不確かさを考慮する動的表現と学習後の回復が、学習型の局所移動を補完的に改善することを示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Learned local navigation in crowded indoor environments is sensitive to how dynamic obstacle motion is represented, while collision-prone behaviour may persist after nominal policy training. We present a risk-aware reinforcement-learning framework that addresses these two issues through an uncertainty-aware Dynamic Uncertainty Grid Map (DUGM) and a modular post-training recovery mechanism. DUGM combines local occupancy, estimated obstacle motion, and motion-estimation uncertainty in a robot-centric representation. After the nominal policy is frozen, a finite-horizon Risk Value Function (RVF) is trained from nominal rollouts and used to trigger a dedicated recovery policy when continued nominal execution is predicted to be collision-prone. Experiments in a held-out NVIDIA Isaac Sim clinical-logistics benchmark show that uncertainty-aware dynamic representation improves nominal navigation over static and deterministic alternatives, while the recovery mechanism further mitigates residual collision-prone behaviour. The complete framework is also deployed directly on a TurtleBot3 without policy fine-tuning, retraining, or site-specific adaptation, retaining the performance trend observed in simulation. These results indicate that uncertainty-aware dynamic representation and post-training recovery provide complementary mechanisms for improving learned local navigation.

著者のコメント

8 pages, 7 figures, 2 tables

arXiv ID: 2609.27338 / 要約の誤りについて