人の作業テンポをロボットの模倣学習に移す
EgoSpeedUp: Transferring Human Manipulation Tempo to Robot Policies
この論文をやさしく読む
ひとことで言うと
人の実演から作業の各段階に合う速さを学び、ロボットの動作を速くする方法です。
何に役立つ?
遅いロボット実演を使った模倣学習で、作業段階ごとの速度を調整する手段になります。
この研究の面白いところ
人とロボットの同じ作業の段階を対応付けて実演の時間配分を変えます。実世界の二つの課題で成功率が平均25ポイント上がり、成功時の所要時間が36.5%減りました。
どこまで分かった?
実証は二つの実世界の操作課題です。異なる作業やロボットでも同じ改善が得られるかは要旨からは分かりません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
模倣学習で訓練したロボットの操作方策は、実演された動作だけでなく、ロボットによる実演の慎重で遅い実行テンポも受け継ぐ。既存の高速化法は元の実演より速く動かせるが、適切な加速の程度を主にロボット側の情報や事前に定めた速度倍率から決めるため、操作の各段階をどの速さで進めるべきかを示す、作業に合った基準をどう得るかは未解決である。 著者らは、人間の操作を時間的な教師情報としてロボットの模倣学習に使うEgoSpeedUpを導入する。人間の実演には、作業に合った段階ごとの操作テンポが自然に表れるという着想に基づく。同じ作業についての遅いロボット実演と人間の実演から、対応する操作段階をそろえ、複数の人間実演から相対的な実行テンポを推定し、その段階ごとのテンポに合わせてロボット実演の時間配分を変更する。その実演を通常の行動模倣学習に使うことで、ロボットが実行できる操作動作を保ちながら、人間の実演に基づいたテンポで作業するよう学習する。実世界の二つの操作課題では、作業の成功率が平均25パーセントポイント上がり、成功した作業の所要時間は36.5%短くなった。この結果は、人間の操作テンポが、より速く信頼性の高いロボット方策を学ぶための有効な時間的基準になることを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Robot manipulation policies trained through imitation learning inherit not only the demonstrated behavior but also the conservative execution tempo of robot demonstrations. Existing acceleration approaches can execute faster than the original demonstrations, but determine the appropriate acceleration primarily from robot-side information or a predefined set of tempo factors, leaving open how to obtain a task-appropriate reference for how fast each manipulation phase should progress. We introduce EgoSpeedUp, a framework that uses human manipulation as temporal supervision for robot imitation learning. Our key insight is that human demonstrations naturally reveal task-appropriate, phase-wise manipulation tempo. Given slow robot demonstrations and human demonstrations of the same task, EgoSpeedUp aligns corresponding manipulation phases, estimates their relative execution tempos from multiple human demonstrations, and transfers the resulting phase-wise tempo by retiming the robot demonstrations. The retimed demonstrations are then used for standard behavior cloning, allowing the robot to retain its executable manipulation behavior while learning to perform it at a human-informed tempo. Across two real-world manipulation tasks, EgoSpeedUp improves the task success rate by an average of 25 percentage points (pp) while reducing successful execution time by 36.5%. These results demonstrate that human manipulation tempo provides an effective temporal reference for learning faster and more reliable robot policies.
著者のコメント
8pages
arXiv ID: 2609.29310 / 要約の誤りについて