arXiv論文メモ
新着一覧
cs.RO / cs.CV · 査読状況未確認

悪路でロボットが受ける揺れや危険を予測する世界モデル

Feeling Terrain Before Crossing: World Models for Off-Road Navigation

E-In Son, Dong-Wook Kim, Ji-Hoon Hwang, Kangsun Lee, Jisung Bae, Jung-Taak Kim and Seung-Woo Seo

この論文をやさしく読む

ひとことで言うと

荒れた地面を進む前に、将来の映像だけでなく滑りや傾きなどロボットが受ける状態も予測して経路を選ぶ方法です。

何に役立つ?

未舗装路で危険な地面を避ける走行計画に役立ちます。自己感覚と失敗リスクを、ロボット自身の経験から人のラベルなしで学習します。

この研究の面白いところ

視覚予測と身体状態の予測を組み合わせ、目標への近さと危険度を別々に評価します。山道のHuskyで機上計画を実行し、比較した端から端までの方策が失敗するコースを完走しました。

どこまで分かった?

実データ、シミュレーション、山道での実機走行を含む成果です。ただし要旨には走行回数や成功率はなく、すべての悪路で安全を保証するものではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

ナビゲーション用の世界モデルは、観測を直接行動へ写すのではなく、候補となる各行動列が生む未来を予測し、最良のものを選ぶという先読みで計画する。予測した風景が十分な代理指標となる都市環境とは異なり、オフロード走行ではロボットと地形の相互作用が決定的である。そのため、予測にはカメラに何が見えるかだけでなく、ロボットが何を感じるかも含める必要がある。しかし、既存の風景中心のモデルは、予定経路に沿ってロボットがどれほど滑り、傾き、揺れるかを予測しない。固有受容情報はこうした運動を直接捉え、入力に使うと物理的な未来の予測を改善する。 本研究は、固有受容情報を条件とし、カメラに見えるものとともにロボットが感じるものを予測する、初のオフロード用ナビゲーション世界モデルFeel-WMを提示する。物理的な未来は、将来の固有受容状態と失敗リスクとして表し、どちらも人手のラベルなしでロボット自身の経験から学習する。計画器は風景と並行して物理的な未来を展開し、分離可能なスコアを用いて、予測した失敗リスクと目標への類似度を比較評価する。 実際のオフロードデータとシミュレーションによる実験では、車輪型と脚型の両方で、Feel-WMが視覚のみのナビゲーション世界モデルを、開ループ計画と閉ループの不整地走行で上回る。山道のHuskyに搭載すると、Feel-WMは機上で計画し、前方の荒れた地面を予測して回避し、エンドツーエンド方策では失敗するコースを走破する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Navigation world models plan by foresight, predicting the future that each candidate action sequence produces and selecting the best, rather than mapping observations to actions directly. Unlike urban settings where a predicted scene is a sufficient proxy, off-road navigation hinges on the robot--terrain interaction, so the prediction must cover not only what the camera will see but what the robot will feel. However, existing scene-focused models do not predict how much the robot will slip, tilt or shake along a planned trajectory. Proprioception captures these dynamics directly and, when used as input, improves the prediction of the physical future. We present Feel-WM, the first off-road navigation world model that conditions on proprioception and predicts what the robot will feel alongside what the camera will see. The physical future takes the form of a future proprioceptive state and a failure risk, both learned from the robot's own experience without human labels. The planner rolls out the physical future alongside the scene and weighs the predicted failure risk against goal similarity in a separable score. Experiments on real off-road data and in simulation demonstrate that Feel-WM outperforms visual-only navigation world models in open-loop planning and closed-loop rough-terrain navigation across wheeled and legged platforms. Deployed on a Husky on mountain trails, Feel-WM plans onboard, predicts rough ground ahead and steers around it, completing courses that an end-to-end policy fails.

著者のコメント

8 pages, 6 figures

arXiv ID: 2609.19863 / 要約の誤りについて