両腕ロボットの役割交換に対応する特徴表現
BiRoAD: Learning Shared and Role-Adaptive Representations for Bimanual Manipulation
この論文をやさしく読む
ひとことで言うと
二つの腕に割り当てる役割が変わっても使えるよう、共通の動きと役割に固有の動きを分けて学習する方法を提案した。
何に役立つ?
両腕ロボットの模倣学習で、実演例が少ない役割の組合せへの対応を改善する用途が考えられる。
この研究の面白いところ
腕の交換で不変な成分と変化する成分を分け、既存の方策へ残差として組み込める点が特徴である。
どこまで分かった?
要旨では複数課題での改善を報告するが、具体的な成功率や実機での条件は示されていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
両腕操作では、二つの腕を協調させながら、場面の幾何、物体の配置、作業内容に応じてそれぞれの役割を変える方策が必要となる。しかし、実演データで役割の分布が偏ると、少ない腕と役割の組合せへ一般化しにくい。また、多くの方策は左右の腕に固定した行動空間を使い、腕の間で機能上の役割を交換したときに行動がどう変わるべきかを明示しない。そこで、共有表現と役割に応じて変わる表現を学ぶBiRoADを提案する。両腕の軌道または行動トークンの特徴を、腕を交換しても変わらない対称成分と、交換に応じて符号が変わる反対称成分に分解する。前者は共通の協調構造を、後者は役割固有の違いを表す。両成分を元の腕の特徴へ残差として組み戻すため、方策の入力や模倣学習の目的を変えたり、人手で役割ラベルを定義したりせずに利用できる。役割分布が均衡した場合と偏った場合を含む複数の両腕操作課題で、対応する元の方策より役割の組合せに対する頑健性が向上し、とくに少数の組合せで改善が目立った。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-20(UTC)
- 最新改訂
- 2026-09-20 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-20 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Bimanual manipulation requires policies that coordinate two arms while adapting their functional roles to scene geometry, object configuration, and task context. Learning such scene-conditioned role adaptation remains challenging, as demonstrations may contain uneven role distributions that limit generalization to underrepresented arm--role configurations. In addition, many bimanual policies predict actions in fixed left- and right-arm action spaces. While this provides a natural parameterization for robot control, it does not explicitly specify how behaviors should transform when functional roles are exchanged across arms. Across different scene initializations, the two arms may follow a similar coordination pattern, but the role-specific behavior assigned to each arm should change with the scene. Therefore, we propose BiRoAD, a Bimanual Role-Adaptive Decomposition framework for learning shared and role-adaptive representations in bimanual policies. Given bimanual trajectory or action-token features, BiRoAD decomposes these features into swap--symmetric and swap--antisymmetric components: the former captures coordination structure invariant to arm exchange, and the latter captures role-specific distinctions that vary consistently with functional role assignment. The two components are then recomposed as residual updates to the original paired arm representations, allowing BiRoAD to serve as a modular feature transformation without changing the policy inputs, imitation-learning objective, or requiring manually defined role labels. Across multiple bimanual manipulation tasks with balanced and imbalanced role distributions, BiRoAD improves robustness across role configurations over corresponding base policies, with notable gains on underrepresented role configurations.
著者のコメント
Accepted at CoRL 2026
arXiv ID: 2609.23445 / 要約の誤りについて