arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

精度不足や別作業のデータをロボットの精密操作学習に活用

Imperfection for Precision: Upcycling Imperfect Data for High-Precision Robotic Manipulation

Hao Wei, Yang Liu, Chao Tang, Shengbao Li, Jiangtao Chen, Jinxuan Zhu, Jiaheng Wang, Hong Yin, Zhaofeng Cao, and Tingguang Li

この論文をやさしく読む

ひとことで言うと

目的の作業だが動きが粗いデータと、別の作業だが動きが精密なデータを、学習の異なる段階で使い分ける方法です。

何に役立つ?

精密ロボット操作に必要な専用の高品質データを集める負担を減らす用途があります。要旨では実機の細かい操作と大まかな操作の両方を評価しています。

この研究の面白いところ

作業の意味を学ぶ段階と、動作の細部を整える段階で、参考にするデータを変えます。単純に全データを混ぜない点が中心です。

どこまで分かった?

31.7と4.2は相対的な改善率ではなくパーセントポイントです。最大改善の実験と、高品質データを置換した際の平均低下の実験は別です。置換しても性能低下が全くないという結果ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

高精度の操作に向けて視覚・言語・行動(VLA)モデルを学習するには、通常、遠隔操作などで得た作業固有の高品質データが必要だが、その収集には時間と費用がかかる。操作精度を損なわずにこの負担を減らすため、本研究では、通常なら捨てられる2種類のデータを有効な資源へ作り替える、単純で効果的な手法ε4P(Imperfection for Precision)を提案する。2種類とは、対象作業の低精度データと、対象とは異なる作業の高精度データである。 ε4Pは、共同学習の全体を通じてこれらの不完全なデータを単純に混ぜるのではなく、フローマッチングの軌道上で各データ源が寄与する場所を制御する。具体的には、大きな雑音の段階で対象作業の低精度データを使い、作業の大局的な文脈を保つ。小さな雑音の段階では、作業が一致しない高精度データを使い、低レベルの行動精度を移す。 1 mm未満の精度を要する高精度作業と、大まかな操作でよい作業の両方について実機ロボットで実験し、本手法が、追加の不完全なデータを有効利用して方策の性能を最大31.7パーセントポイント改善できること、および同量の作業固有の高品質データを置き換えても、性能の低下を平均4.2パーセントポイントにとどめられることを示す。全体としてε4Pは、異質で不完全なデータを体系的に再利用し、高価な作業固有の高品質データへの依存を減らす、拡張可能な高精度操作の方式を示す。詳細はhttps://varepsilon4p.github.io/で公開されている。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Training vision-language-action (VLA) models for high-precision manipulation typically requires task-specific, high-quality data (e.g., teleoperation), which is slow and expensive to collect. To reduce this burden without compromising manipulation precision, we propose $\varepsilon$4P (Imperfection for Precision), a simple yet effective method that "upcycles" two otherwise discarded data sources: (1) low-precision data from the target task and (2) high-precision data from mismatched tasks. Rather than naively mixing these imperfect data sources throughout co-training, $\varepsilon$4P controls where each source contributes along the flow-matching trajectory. Specifically, low-precision, target-task data is used at high noise to preserve high-level task context and high-precision, task-mismatched data is used at low noise to transfer low-level action precision. Through real-robot experiments on both sub-millimeter, high-precision tasks and coarse-grained tasks, we demonstrate that the proposed method (1) effectively leverages additional imperfect data to improve policy performance by up to 31.7 percentage points, and (2) can replace an equal amount of task-specific, high-quality data with an average performance drop of only 4.2 percentage points. Overall, $\varepsilon$4P points toward a scalable paradigm for high-precision manipulation, in which heterogeneous, imperfect data can be systematically repurposed to reduce reliance on costly task-specific, high-quality data. More details are available at https://varepsilon4p.github.io/.

著者のコメント

9 pages, 5 figures

arXiv ID: 2609.26672 / 要約の誤りについて