arXiv論文メモ
新着一覧
cs.RO / cs.AI / cs.CV / cs.LG · 査読状況未確認

地図上の位置を使って移動マニピュレーターを制御するMAVP

MAVP: Map-Aware Visuomotor Policies for Mobile Manipulation

Jinhe Tang, Ruixiao Dai, Weiming Zhi

この論文をやさしく読む

ひとことで言うと

ロボットの台車の目標位置を地図上で予測し、フィードバックでずれを直す操作方策。

何に役立つ?

移動と腕の操作を組み合わせるロボットで、位置ずれによる失敗を減らす方法として参考になる。

この研究の面白いところ

実演の軌跡を共通の地図座標にそろえ、台車・腕・グリッパーを同時に予測している。

どこまで分かった?

成功率の改善は実世界の六課題、三種類の方策での比較結果であり、要旨に他の環境での性能は示されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

移動しながら物体を操作するには、台車と腕の動きを協調させ、空間内の位置を正確に保つ必要がある。しかし、実演から学習した方策は意図した台車の動きを安定して実行できず、位置ずれから操作に失敗することがある。本研究は、台車の目標姿勢を明示的に予測し、位置推定のフィードバックを使って追従することで、実行の信頼性を改善するMap-Aware Visuomotor Policies(MAVP)を提案する。MAVPは遠隔操作による実演から静的な地図を再構成し、実演中の台車の軌跡を共通の地図座標系で表すことで、複数の実演に一貫した空間的な教師信号を与える。実行時には、方策がRGB画像、関節状態、地図座標系での現在の台車姿勢を受け取り、台車の目標姿勢、腕の動作、グリッパーの動作を同時に予測する。低水準の制御器は、先行的な動作と姿勢誤差のフィードバックを使って目標姿勢を追跡し、実行時のずれを修正する。また学習中に姿勢ノイズを加え、方策への姿勢入力の誤差に対する頑健性を高める。実世界の六つの操作課題と三種類の方策で、MAVPは地図に固定しない速度制御よりすべての課題で高い成功率を達成した。動画と追加の結果も公開している。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Successful mobile manipulation requires coordinated base and arm motion while maintaining accurate spatial positioning. However, demonstration-trained policies can struggle to realise the intended base motion reliably, leading to spatial misalignment and subsequent manipulation failures. We present MAVP (Map-Aware Visuomotor Policies), a framework that improves execution reliability by predicting explicit base-pose targets and tracking them using localisation feedback. MAVP reconstructs a static map from teleoperated demonstrations and expresses demonstrated base trajectories in a shared map frame, providing consistent spatial supervision across demonstrations. At execution time, the policy receives RGB observations, joint states, and the robot's current map-frame base pose, and jointly predicts target base poses, arm actions, and gripper actions. A low-level controller tracks the predicted base targets using feedforward motion and pose error feedback, enabling correction of execution deviations. We additionally use pose-noise augmentation during training to improve robustness to errors in the policy's pose input. Across six real-world manipulation tasks and three policy families, MAVP achieves higher task success rates than unanchored velocity control in all tasks. Videos and additional results are available at https://123qwedsa123.github.io/mavp/.

arXiv ID: 2609.26378 / 要約の誤りについて