arXiv論文メモ
新着一覧
cs.RO / cs.AI · 査読状況未確認

対象物を追跡しながら長い手順を実行するロボット制御

MaskHarness-WAM: Instance-Grounded Harnessing for Long-Horizon Robot Manipulation

Zitai Huang, Taiyi Su, Jian Zhu, Jianjun Zhang, Chong Ma, Tianbin Liu, Weiyi Lu, Yi Xu, Hanli Wang

この論文をやさしく読む

ひとことで言うと

見た目が同じ物体を決まった順で扱うロボットに、対象の追跡と作業の区切りを管理する仕組みを加えています。

何に役立つ?

一つの短い操作ができる方策を、複数物体の連続した作業へ広げるために役立ちます。

この研究の面白いところ

対象マスクで計画と操作を結び、各小課題の境界で環境を見直して新たなマスクを生成・確認します。完了が確かめられた段階で次の物体へ移ります。

どこまで分かった?

実機で短い時間範囲の方策より大きく改善したと報告しています。要旨には成功率、物体数、作業の種類の詳細はなく、任意の長時間作業での信頼性まで確認したものではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

長い時間にわたるロボット操作では、局所的な視覚運動制御が安定しているだけでなく、実行を通じた継続的な対象追跡と、信頼できるタスク進捗評価が必要である。この課題は、外見が同じ複数の物体を決められた順序で操作しなければならない場合に特に重要となる。そのような場面では、限られた時間範囲を扱う操作方策だけに頼っても、どの個体を操作すべきか、いつ次の段階へ移るべきかを十分に判断できないことが多い。 この課題に対し、個々の対象物に対応付けて長時間の操作を統括する仕組み、MaskHarness-WAMを提案する。提案システムは対象マスクを介して上位のタスク計画と下位の操作方策を結び付け、視覚フィードバックを使ってサブタスクの実行順を調整し、継続的に実行する。各サブタスクは異なる対象個体に対応するため、下位の方策には、サブタスクが切り替わるたびに、更新された場面に応じた新しい初期対象マスクが必要になる。 この統括機構は環境を継続的に再観測し、サブタスクの境界で対象マスクを生成・検証することで、下位の方策に与える個体単位の空間条件を更新する。さらに、各サブタスクの検証済み完了状態に応じて対象個体を切り替え、操作過程を進める。実機ロボットでの実験では、複数物体を順番に操作するタスクで、MaskHarness-WAMが短い時間範囲の方策を大幅に上回る。これは、局所的な操作技能を、信頼できる長時間の実行へと拡張する上での有効性を示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Long-horizon robot manipulation requires not only stable local visuomotor control, but also continuous target tracking and reliable task progress assessment throughout execution. This challenge becomes particularly critical when multiple objects share identical appearances and must be manipulated in a prescribed order. In such scenarios, relying solely on a limited-horizon manipulation policy is often insufficient to determine which instance should be operated on and when the task should transition to the next stage. To address this challenge, we propose MaskHarness-WAM, an instance-grounded harness for long-horizon manipulation. The proposed system connects high-level task planning with low-level manipulation policies through target masks, while leveraging visual feedback for subtask scheduling and continuous execution. Since each subtask corresponds to a different target instance, the low-level policy requires a newly established initial target mask under the updated scene at each subtask transition. The harness continuously re-observes the environment, generates, and verifies the target mask at subtask boundaries, thereby updating the instance-level spatial condition provided to the low-level policy. Furthermore, the system advances the manipulation process by switching target instances according to the verified completion status of each subtask. Experiments on a real robot platform demonstrate that MaskHarness-WAM substantially outperforms limited-horizon policies on sequential multi-object manipulation, showing its effectiveness in extending local manipulation skills to reliable long-horizon execution.

arXiv ID: 2609.19974 / 要約の誤りについて