人の動作を使ってヒューマノイドの自己位置推定を学習
PRIMO: Prior-Informed Odometry from Human-Motion Tracking for Humanoid Robots
この論文をやさしく読む
ひとことで言うと
ロボット自身のセンサーから移動を推定するモデルを、人の多様な動作をまねるシミュレーションで学び、物理的な事前知識で実機への移行を助ける方法です。
何に役立つ?
制御方策が変わっても使えるヒューマノイドの移動推定を目指します。実機でも比較手法に対する誤差低減が報告されており、シミュレーションだけの評価ではありません。
この研究の面白いところ
学習データを配備予定の制御方策に依存させない工夫と、予測モデルに物理・対称性を取り込む工夫を組み合わせています。生のセンサー文脈を残す別経路も使います。
どこまで分かった?
31.6〜61.7%は実機での外部比較、86.8〜94.6%は反対方策のシミュレーション、69.2〜81.7%は実機動作で学習データを比較した値です。比較条件が違うため、一つの改善率としてまとめることはできません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
シミュレーションで学習するヒューマノイドの自己受容感覚に基づくオドメトリには、二つの移行上の課題がある。実配備を想定した特定のロボット制御方策から生成する訓練軌跡は、動きの範囲が限られる。また、シミュレーションと実世界の不一致により、制約のない予測は信頼できなくなり得る。本研究では、人動作追従からの事前知識を用いたオドメトリPRIMOにより、両方に対処する。データ面では、ロボット向けに変換した多様な人の動作をシミュレーション内でヒューマノイドに追従させ、オドメトリの教師情報を生成する。これにより教師情報を配備時の方策から切り離し、学習時の動作分布を広げる。モデル面では、事前知識を用いる推定器が、物理と対称性に基づく事前知識で速度・回転予測に構造を与え、さらに大まかな生の文脈情報を通す経路によって、符号化した特徴と並行してセンサーの文脈を保持することで、実世界への汎化を強める。統一した実機評価手順の下で、PRIMOは各領域・指標の比較における評価済み外部比較手法の最良のものに対し、平均誤差を31.6〜61.7%減らす。二つの歩行制御方策の改訂版にまたがる評価では、方策ごとの特化モデルは互いの方策で対称的な性能の逆転を示す一方、Tracking-Locomotionによる学習は、反対側の方策におけるシミュレーション平均誤差を86.8〜94.6%減らす。実機での動的な動作では、Tracking-Locomotionによる学習は、両方の配備方策を合わせたデータで学習する場合に比べ、平均誤差を69.2〜81.7%減らす。検証した動作構成全体にわたり、事前知識を用いる推定器は、シミュレーションと実機の双方で、制約なしの対応モデルより平均軌跡誤差を一貫して小さくする。コードは https://github.com/Agibot-Spatial-Intelligence/PRIMO で公開されている。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-20(UTC)
- 最新改訂
- 2026-09-20 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-20 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Simulation-trained humanoid proprioceptive odometry faces two transfer challenges: training trajectories generated by specific robot control policies intended for deployment cover only a limited range of motions, while sim-to-real mismatch can make unconstrained predictions unreliable. We address both with Prior-Informed Odometry from Human-Motion Tracking (PRIMO). On the data side, we generate odometry supervision by having the humanoid track diverse retargeted human motions in simulation, decoupling supervision from the deployment policies and broadening the training motion distribution. On the model side, a Prior-Informed estimator uses physics- and symmetry-informed priors to structure velocity and rotation prediction and a coarse raw-context pathway to preserve sensor context alongside encoded features, thereby strengthening sim-to-real generalization. Under a unified real-robot protocol, PRIMO reduces mean error by 31.6%-61.7% relative to the strongest evaluated external baseline in each domain-metric comparison. Across two locomotion-policy revisions, policy specialists exhibit symmetric crossover, whereas Tracking-Locomotion training reduces mean opposite-policy simulation error by 86.8%-94.6%. On real dynamic motion, Tracking-Locomotion training reduces mean error by 69.2%-81.7% relative to training on the union of both deployment policies. Across the tested motion compositions, the Prior-Informed estimator consistently lowers mean trajectory errors relative to its Unconstrained counterpart in both simulation and real-robot evaluation. Code is available at https://github.com/Agibot-Spatial-Intelligence/PRIMO.
著者のコメント
8 pages, 6 figures and 4 tables, Under review
arXiv ID: 2609.23610 / 要約の誤りについて