arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

壊れやすい物体の把持を触覚反射の実演から学ぶ

What is the Better Curriculum: Controller-Shaped Grasping Behavior for Contact Force-Sensitive Manipulation

Ziyan Feng, Zizhao Yuan, Yulong Fu, Yuxin He, Zhiyuan Zhang, Zhengjie Zhang, Jinni Zhou, Renjing Xu, Qiang Nie

この論文をやさしく読む

ひとことで言うと

壊れやすい物体をつかむ実演を、触覚反射コントローラーで安定させてから、触覚なしのロボット方策に学ばせた。

何に役立つ?

考えられる用途は、狭い接触力の範囲を守る必要がある把持の訓練データ作成である。

この研究の面白いところ

標準のプラスチックカップでは95%の安定把持を示したが、外乱下では学習した方策だけで45%失敗し、運用時の反射制御がなお有効だった。

どこまで分かった?

紙コップの結果は探索的な傾向として述べられている。外乱への頑健性は触覚なしの方策だけでは達成されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

1 N未満の接触力でも取り返しのつかない損傷を受けるほど壊れやすい物体を、ロボットはどのように操作すべきだろうか。既存の視覚・触覚方策学習は触覚を方策への追加入力とみなすことが多い。しかし直接接触する力に敏感な操作では、障害はもっと早いデータ収集段階で生じ得る。手動のグリッパー操作は遅れが大きく粒度も粗いため、安定した把持に必要な狭い力の範囲を確実に保つことが難しい。そこで著者らは、25 Hzで動く決定的な触覚反射コントローラーを収集時の教師として使い、コントローラーが形作った把持動作の実演データを作り、触覚なしの方策を学習させる。Action Chunking with Transformers(ACT)では、反射制御により形作った実演から学んだ方策が教師の把持の力の変化を再現し、標準のプラスチックカップ課題で安定把持率95%を達成して、映像を確認した手動実演より大きく優れた。同じ介入はπ0.5の訓練分布内での安定性も改善し、未学習の紙コップの変種でも探索的な良い傾向を示した。しかし外乱を無作為に与えると、反射制御データで学んだπ0.5方策は方策だけで動かす試行の45%で失敗した一方、運用時にも反射制御を仲裁に使うとすべての把持を維持した。これらの結果は、触覚フィードバックを方策に組み込む代わりに、収集時の教師として実演の把持動作を形作る役割を示す。同時に、外乱への対処はなおコントローラーに依存し、触覚を使わない方策の限界も示している。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

How should a robot learn to manipulate objects so fragile that sub-Newton contact forces can cause irreversible damage? Existing visuo-tactile policy learning typically treats tactile sensing as an additional policy input. In direct-contact force-sensitive manipulation, however, the bottleneck can arise earlier, during data collection: manual gripper control is too delayed and coarse-grained to reliably maintain the narrow force range required for stable grasping. We therefore use a deterministic 25 Hz tactile reflex controller as a collection-time teacher, producing demonstrations with controller-shaped grasping behavior for tactile-free policy learning. On Action Chunking with Transformers (ACT), policies trained from reflex-shaped demonstrations recover the teacher's grasping profile and achieve 95% stable grasps on the nominal plastic-cup task, substantially outperforming visually screened manual demonstrations. The same intervention improves in-distribution stability on $\pi_{0.5}$ and shows a favorable exploratory trend on an unseen paper-cup variant. Under randomized external disturbance, however, the reflex-data $\pi_{0.5}$ policy still fails in 45% of policy-only trials, whereas a deployment-time reflex arbiter retains all grasps. These results reveal a new role for tactile feedback in force-sensitive manipulation: rather than integrating tactile into the policy, we use it as a collection-time teacher that shapes grasping behavior in demonstrations for policy learning, while disturbance rejection remains controller-dependent, revealing the boundary of tactile-free policy.

著者のコメント

22 pages, 6 figures. Project page: https://shayfeng.github.io/better-curriculum/

arXiv ID: 2609.25887 / 要約の誤りについて