未知の物体を押しながら動き方を学ぶMetaPusher
MetaPusher: Meta Learning and Planning for Nonprehensile Manipulation of Unseen Objects with Rapid Online Adaption
この論文をやさしく読む
ひとことで言うと
初めての物体を押しながら性質を学び、それに合わせて先の動作計画を修正する仕組みです。
何に役立つ?
考えられる用途は、摩擦などを事前に測れない物体をロボットで押して配置する作業です。未知物体の評価で予測誤差と成功率の改善を報告しています。
この研究の面白いところ
学習で物体のモデルが変わるたび、探索木を使い回して計画も直します。作業と適応を同時に進める点が特徴です。
どこまで分かった?
最大20%の改善が相対値か百分率ポイントかは要旨で明示されません。任意の未知物体に対する成功保証ではありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
それまで扱ったことのない物体の操作は、今なお難しい。その運動は、摩擦や質量分布のように知覚だけからは推定できない潜在的な物理特性に依存するためである。複数の物体についての過去の経験は未知物体の動力学の初期推定を与えられるが、不確実さが残り、シミュレーションから実世界への移行でさらに悪化し得る。相互作用を通して動力学を適応させれば推定を徐々に改善できる一方、モデルの更新によって計画済み軌道が使えなくなることがある。効率よく操作を成功させるには、迅速な動力学の適応と、その変化を取り込める計画戦略の両方が必要である。 本研究では、物体ごとの事前の相互作用なしに未知物体を把持せず操作するための、メタ学習と適応的計画の枠組みMetaPusherを導入する。メタ学習した動力学モデルはタスク実行中の相互作用から素早く適応し、適応的な運動学・動力学計画器は既存の探索木を再利用・改良して長い時間範囲の計画を更新する。この結合により、独立したデータ収集段階なしに操作と適応を進められる。未知物体を用いたシミュレーションとシミュレーションから実世界への移行場面で評価し、ファインチューニング手法、能動学習手法、MPPIに基づく制御、および強化学習方策と比較する。予測誤差を低減し、タスク成功率を最大20%向上させた。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-17(UTC)
- 最新改訂
- 2026-09-17 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-17 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Manipulating previously unseen objects remains challenging, as their dynamics depend on latent physical properties, such as friction and mass distribution, that cannot be inferred from perception alone. Prior experience across objects can provide an initial estimate of unseen object dynamics, but this estimate remains uncertain and can degrade further during sim-to-real transfer. Adapting the dynamics through interaction can progressively refine the estimation, however, updating the model may invalidate the planned trajectory. Successful and efficient manipulation therefore requires both rapid dynamics adaptation and a planning strategy that can incorporate this evolution. In this work, we introduce MetaPusher, a meta-learning and adaptive planning framework for nonprehensile manipulation of unseen objects without prior object-specific interactions. A meta-learned dynamics model rapidly adapts from interactions during task execution, while an adaptive kinodynamic planner updates long-horizon plans by reusing and refining its existing search tree. This coupling enables manipulation and adaptation without a separate data collection phase. We evaluate MetaPusher on unseen objects in simulation and in sim-to-real scenarios, comparing against fine-tuning and active learning methods, MPPI-based control, and a reinforcement learning policy. It achieves lower prediction error and improves task success rate by up to 20%.
arXiv ID: 2609.21122 / 要約の誤りについて