arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

ロボットの動作空間の形を保って高速に行動を生成

Faster Visuomotor Policy Learning on Action Manifolds via Riemannian MeanFlow

S. Talha Bukhari, Austin Garrett, Yi Wei, Ruiqi Ni, Zachary Kingston, Aniket Bera

この論文をやさしく読む

ひとことで言うと

ロボットの動作空間の形を守りながら、少ないモデル計算で次の動作列を作る方法です。

何に役立つ?

素早い動作生成が必要なロボット制御で、拡散・フロー型方策の計算負担を下げる設計に役立つと考えられます。

この研究の面白いところ

ネットワーク評価1回でも動作多様体上の列を作れ、複数のベンチマークと実ロボット操作で試しています。

どこまで分かった?

要旨は従来法と競争力がありサンプリング費用が低いと述べていますが、具体的な速度倍率は示していません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

視覚運動方策は、未加工の感覚情報からロボットの動作列へ直接写す方法を学習する。拡散モデルやフローマッチングに基づく方策は、動作列の多峰的な分布を端から端まで捉えられる。しかし、その表現力には、動作生成時に学習したベクトル場を複数段階で数値積分する費用が伴い、ロボットに必要な高速制御を妨げる。また、ロボットの動作列は通常、滑らかで微分可能な多様体上に定義されるため、方策は動作空間の本来の幾何を尊重しなければならない。本研究は、ロボットの動作多様体上で、確率経路の条件付きフロー写像を学習するRiemannian MeanFlow Policy(RMFP)を提示する。定式化には、データに基づくリーマン条件付きフローマッチングを基点とする、フロー写像の整合性に関する目的関数を用いる。この整合条件は学習時に安定で、モデルを有限時間の移送に制約するため、ネットワークの関数評価をわずか1回にしても多様体上の動作列を生成できる。球面上のLASA、Push-T、RobomimicのTool HangとTransport、および多様体に制約された動作生成を行うFranka Kitchenで結果を示し、従来手法と競争力のある性能を、より少ないサンプリング費用で得た。実世界のロボット操作課題にも適用し、不完全なセンサー測定の下でも現実の環境で動作を高速に生成できることを示した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Visuomotor policies learn a direct map from raw sensory observations to robot action sequences. Policies based on Diffusion and Flow Matching capture the multimodal distribution over action sequences in an end-to-end manner. This expressivity comes at the cost of multi-step numerical integration of the learned vector field for action generation, which can be expensive and time-consuming, impeding fast control rates required in robotics applications. Furthermore, robot action sequences are usually defined on a smooth, differentiable manifold, requiring that the learned policy respects the intrinsic geometry of the robot's action space. Here, we present Riemannian MeanFlow Policy (RMFP), which learns the conditioned flow map of the probability path on the robot action manifold. Our formulation employs a flow map consistency objective grounded in the data by a Riemannian Conditional Flow Matching anchor. The flow map consistency condition is stable to train and constrains the learned model to finite-time transport, which yields on-manifold action sequence generation with as few as one network function evaluation. We present results on the spherical LASA and Push-T benchmarks, on the Tool Hang and Transport tasks of the Robomimic suite, and on the Franka Kitchen task with manifold-constrained action generation, and demonstrate that RMFP attains performance competitive with prior work at a lower sampling cost. We also employ RMFP on a real-world robotic manipulation task to demonstrate fast action generation under imperfect sensor measurements in the physical world.

arXiv ID: 2609.30127 / 要約の誤りについて