物体が見えていない場面での作業と動作の計画
Search, Ground, Plan: Functional Sufficiency for Task and Motion Planning under Incomplete Scene Knowledge
この論文をやさしく読む
ひとことで言うと
作業に必要な物体が本当にあるか確かめてから、ロボットの動作を計画する研究。
何に役立つ?
見えていない物体がある場面で、実行不能な計画を避ける設計の参考になる。
この研究の面白いところ
機能的な役割を全て実物に割り当てられるまで探索を続ける点。
どこまで分かった?
成功率54.0%は指定された32場面・200試行での値であり、なお失敗も残っている。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
基盤モデルにより、言葉と画像で指定された物体操作へ作業・動作計画を広げられるようになった。しかし場面についての知識が不完全だと、作業に必要なものを理解することと、実際の場面で実現できると判断することの間に重要な隔たりがある。本研究は、作業完了に必要な場面の物体を探し、機能的な役割を適切な実物へ結びつけ、全役割を同時に割り当てて機能上十分であると確かめた後にだけ計画する、基盤モデル型のGRAB-TAMPを導入する。作業を機能的な役割、関係、割当制約で表し、要件が解決していない間は場面を少しずつ調べ、候補物体を意味、形状、関係の検査で確認する。台所、居間、作業場にまたがる32種類の場面で評価した。実行可能な200試行では、端から端までの成功率54.0%、計画目標の達成範囲67.3%だった。同じ実行設定で三つの基盤モデル型作業・動作計画の枠組みと比べると、端から端までの成功率は比較手法の平均より25.7ポイント高かった。実装と評価コードの公開先として https://github.com/Narendhiranv04/GRAB-TAMP を示している。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-19(UTC)
- 最新改訂
- 2026-09-19 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-19 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Foundation models (FMs) have expanded task and motion planning (TAMP) to manipulation problems specified through language and visual observations. However, incomplete scene knowledge leaves a critical gap between understanding what the task requires and knowing whether the physical scene can actually realize it. We introduce GRAB-TAMP, an FM-based TAMP framework that searches for scene entities required for task completion, grounds functional roles to valid physical objects, and plans only after a complete joint assignment establishes functional sufficiency. We represent the task through functional roles, relations, and assignment constraints, and incrementally inspect the scene while requirements remain unresolved, verifying candidate objects through semantic, geometric, and relational checks. We evaluate GRAB-TAMP across 32 scene variants spanning Kitchen, Living Room, and Workshop domains. Across 200 feasible trials, our approach achieves 54.0% end-to-end success with 67.3% plan goal coverage. Compared with three FM-based TAMP frameworks under the same execution setting, GRAB-TAMP improves end-to-end success by 25.7 percentage points over the mean baseline. Implementation and evaluation code: https://github.com/Narendhiranv04/GRAB-TAMP
著者のコメント
8 pages, 4 figures, 4 tables
arXiv ID: 2609.23113 / 要約の誤りについて