指先センサーを自己位置感覚に対応付ける模倣学習
Self-Supervised Anchoring of Fingertip Sensing to Proprioception and Proactive Actions for Robot Imitation Learning
この論文をやさしく読む
ひとことで言うと
ロボットの指先の近接センサーと触覚センサーを、動作や自己位置の情報に対応付けて模倣学習に使う研究です。
何に役立つ?
考えられる用途は、カメラから見えにくい接触前後の状態を使った物体操作です。実世界の操作課題で平均成功率の改善を示しました。
この研究の面白いところ
接触前は近接、接触後は触覚が有用という違いに合わせ、各センサーを固有感覚と行動に独立に対応付けます。
どこまで分かった?
要旨は実験した操作課題での改善を述べていますが、具体的な成功率や未評価のロボットでの性能は示していません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
ロボットの模倣学習は外部カメラに頼ることが多い。しかし、指先付近では遮蔽や時間分解能の不足により、物体との距離、接触の始まり、把持状態などの局所的な相互作用を観測しにくい。本研究は、圧力を測る触覚センサーと反射式の近接センサー、事前学習済みのセンサー符号化器を使い、補完的な指先の感覚を模倣学習へ効果的に取り込む方法を調べる。近接情報は接触前、触覚情報は接触後に有用である。ただし、これらを単純に方策へ加えても性能が安定して向上せず、映像だけの方策を下回る場合もある。少数の実演から、局面によって重要度が変わるまばらな信号を利用するのは難しいことがうかがえる。 そこで、固有感覚に基づく事前学習法PROPRAを提案する。各指先センサーの履歴を、固有感覚と行動の時間区間にそれぞれ独立に対応付ける。これにより、常に得られる感覚運動の基準を用意し、各センサーが有用な局面で個別に位置付けられる。実世界の物体操作課題での実験では、映像だけの方策、および画像を基準にした事前学習より平均成功率が向上した。表現の解析からも、接触前の状態に関するより豊かな情報を保持し、補完的な指先感覚を効果的に使えることが示された。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Robotic imitation learning often relies on external cameras, yet local interaction cues such as object proximity, contact onset, and grasp state are difficult to observe near the fingertips because of occlusion and limited temporal resolution. We study how to effectively incorporate complementary fingertip sensing into imitation learning using pressure-sensitive tactile and reflective proximity sensors, along with pretrained sensor encoders. The two modalities provide information at different manipulation phases: proximity sensing is informative before contact, whereas tactile sensing becomes informative after contact. However, naively adding these signals to a policy does not consistently improve performance and can even underperform vision-only policies, suggesting that sparse, phase-dependent sensor signals are difficult to exploit from limited demonstrations. We therefore propose a proprioception-anchored pretraining method, PROprioceptive-and-PRoactive Anchoring (PROPRA), which independently aligns each fingertip sensor history with proprioceptive and action segments. This provides a continuously available sensorimotor reference, allowing each sensor to be aligned independently during its informative phases. Experiments on real-world manipulation tasks show that our pretraining method improves average success rates over vision-only policies and image-anchored pretraining baselines. Representation analysis further shows that it preserves richer information about pre-contact states, enabling more effective use of complementary fingertip sensing. Please refer to our project page: https://tomohiromotoda.github.io/nia.propra/
著者のコメント
Project page is available at https://tomohiromotoda.github.io/nia.propra/
arXiv ID: 2609.29822 / 要約の誤りについて