ロープを動かした履歴から操作を調整するRopeFormer
RopeFormer: Cross-Trial Adaptation from Interaction History for Dynamic Rope Manipulation
この論文をやさしく読む
ひとことで言うと
ロボットがロープを動かした過去の反応を覚え、次の操作に生かす方法を試した。
何に役立つ?
物性が事前に分からないロープなどの変形物を操作する制御に役立つ可能性がある。
この研究の面白いところ
モデルを再学習せず履歴だけを保持し、実物ロープで試行を重ねるほど複数の成績が改善した。
どこまで分かった?
改善の程度はロープの動きや観測条件に依存する。実機結果はUnitree H1-2と要旨に記された課題での評価である。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
ロープの動的操作は未知の物体の動きに敏感であり、同じロボット動作でもロープによって反応が大きく異なる。一方、関係する物性値を明示的に特定することは難しい。RopeFormerは、以前の試行での行動と反応の履歴を次の制御の文脈に使う枠組みである。方策の重みは固定したまま試行間の履歴を保持し、オンラインでロープのパラメータを明示的に推定する必要はない。片腕の持続回転、両腕の回転、一時的な鞭打ち動作について条件を合わせたシミュレーション評価では、同じ学習済みモデルの履歴を毎回消す場合に比べ、履歴を保つと次の制御が改善した。効果の大きさはロープの動きや観測条件によって異なる。さらに、学習時に見ていない実物のロープを使い、重みを固定した方策をUnitree H1-2へ導入した。試行T1からT3にかけて、Rope Swingの目標到達時間は30.9%、Rope Twirlは33.9%短くなり、Rope Whipの目標命中数の平均は3回中0.2回から2.3回へ増えた。過去の相互作用が変形物の動的操作に有効な制御文脈となることを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-20(UTC)
- 最新改訂
- 2026-09-20 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-20 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Dynamic rope manipulation is highly sensitive to unknown object dynamics: the same robot motion can produce substantially different responses across ropes, while explicitly identifying the relevant physical properties is difficult. We present RopeFormer, a history-conditioned framework that uses prior task interaction as context for subsequent control. The policy retains cross-trial action-response history while keeping its weights fixed and requires no explicit online rope-parameter estimation. In matched simulation evaluations across sustained single-arm rotation, bimanual rotation, and transient whipping, retaining context improves subsequent control relative to resetting the same checkpoint, with the benefit varying across rope dynamics and observation settings. We further deploy the frozen policies on a Unitree H1-2 with previously unseen physical ropes. From T1 to T3, target-acquisition time decreases by 30.9% for Rope Swing and 33.9% for Rope Twirl, while mean Rope Whip target hits increase from 0.2 to 2.3 out of three. These results show that prior interaction can provide effective control context for dynamic deformable-object manipulation. Robot videos, code, and data are available at https://ropeformer.github.io/.
arXiv ID: 2609.23432 / 要約の誤りについて