物理挙動と見た目を再現するロボット操作データ生成
PhyVisGen: Physically and Visually High-Fidelity Robotic Manipulation Data Generation
この論文をやさしく読む
ひとことで言うと
ロボット操作の学習用データを、柔らかい接触と実際の場面の見た目を再現したシミュレーションから生成する方法。
何に役立つ?
実機の実演データを集めにくい操作課題で、合成データから方策を学ぶ方法として役立つ可能性がある。要旨では五つの実ロボット課題での成功率が報告されている。
この研究の面白いところ
柔らかいグリッパーの接触を扱う物理計算と、実場面の再構成を使った画像生成を組み合わせ、合成データだけで学習した方策を実機で評価している。
どこまで分かった?
実機での成功率65〜95%は五つの課題に対する結果であり、すべての操作やロボットに一般化するという記載はない。各課題の条件や失敗の内訳は要旨に記載されていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
頑健な視覚運動方策の学習には大規模な操作実演データが欠かせないが、現実世界での収集は費用がかかり、規模を拡大しにくい。シミュレーションは代替手段となるものの、物理挙動や見た目の差によって、特に柔らかいグリッパーを使う操作では合成データの転用が難しくなる。 本論文では、物理挙動と見た目の忠実度を高め、大規模なロボット操作データを生成する枠組みPhyVisGenを提案する。物理面では、Incremental Potential Contact(IPC)に基づくロボットアームとグリッパーの結合方法を導入し、操作軌道の全体にわたり柔らかい接触を高い忠実度で表現する。視覚面では、実際の場面の再構成とリアルタイムのパストレーシングを組み合わせ、撮影した場面の外観を保ちながら写実的な観測画像を生成する。 定量評価では、PhyVisGenの物理的・視覚的な忠実度が示された。合成した操作実演だけで学習した方策は、実機の実演データも方策の追加調整も使わず、実ロボットの五つの課題で65〜95%の成功率を達成した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-22(UTC)
- 最新改訂
- 2026-09-22 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-22 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Large-scale manipulation demonstrations are essential for learning robust visuomotor policies, yet real-world data collection is expensive and difficult to scale. Simulation offers a promising alternative, but physical and visual discrepancies can limit the transferability of synthetic data, particularly for manipulation with soft grippers. We present PhyVisGen, a physically and visually high-fidelity framework for scalable robotic manipulation data generation. On the physical side, PhyVisGen introduces an arm-gripper coupling method based on the Incremental Potential Contact (IPC), enabling high-fidelity soft contact throughout complete manipulation trajectories. On the visual side, it combines real-scene reconstruction with real-time path tracing to generate visually realistic observations while preserving captured scene appearance. Quantitative evaluations demonstrate the physical and visual fidelity of PhyVisGen. Policies trained exclusively on synthetic manipulation demonstrations achieve 65-95% success across five real-robot tasks, without real-robot demonstration data or policy fine-tuning.
著者のコメント
8 pages, 5 figures. Under review
arXiv ID: 2609.25653 / 要約の誤りについて