arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

シミュレーションだけで学習したロボットの精密挿入

InsertAnything: Generalizable Contact-Rich Precision Insertion from Simulation to Reality

Zhenghua Ma, Xinpan Meng, Zeyu Liu, Muyuan Ma, Hengdi Zhang, Houcheng Li, Long Cheng

この論文をやさしく読む

ひとことで言うと

ロボットが穴に部品を精密に差し込む動作を、シミュレーションだけで学習し、実機でそのまま使う研究です。位置の見積もりがずれても、指先の力を頼りに修正します。

何に役立つ?

考えられる用途は、部品や穴の形が変わる組立工程で、実機の実演収集や再訓練を減らすことです。実験では未経験の8課題で95.0%の成功を報告しています。

この研究の面白いところ

六角形の挿入だけで訓練した方策を、異なる実物の課題へ転用しています。学習時にはシミュレーション、評価時には実機という区別が明確な結果です。

どこまで分かった?

20/20は人が関与するベンチマークのプロトコル下での成績で、挿入動作が自律的という意味です。工程全体が無人という主張ではありません。最小0.02 mmは公称隙間です。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

接触を多く伴う精密な挿入は、ロボット組立に不可欠な操作技能である。隙間が狭いと位置合わせ誤差に敏感になり、衝突や詰まりが起きやすくなる。さらに部品ごとの形状や隙間の違いが方策の再利用を難しくする。本研究では、実世界の実演や方策の微調整を使わず直接配備するため、挿入方策を完全にシミュレーションで訓練する強化学習の枠組みを提示する。 目標姿勢とコンパクトな3次元の指先力フィードバックを組み合わせることで、方策は推定した穴の位置に誤差があっても、位置合わせを探索し動作を修正することを学ぶ。分離したゲート付き報酬によって位置合わせと挿入を調整する。力信号の平滑化と状態に依存しない標準偏差により、学習過程を安定させる。得られた方策は、公称隙間が最小0.02 mmの複数の穴形状で実世界の挿入を行い、穴位置に誤差がある状況で成功率を改善しつつ接触力の最大値を減らす。異なる隙間と形状にまたがる評価も、方策の汎化を確認した。 本システムは、ManipulationNetの穴へのペグ挿入ベンチマークにおいて、人がループに関与するプロトコルの下で初めて20/20の満点を達成し、挿入動作自体は完全に自律的だった。シミュレーション上の六角形挿入課題だけで訓練した単一の方策は、未経験の実世界の挿入課題8種で全体成功率95.0%を達成した。これらの結果は、シミュレーションだけの学習によって、実世界の課題に直接配備し再利用できる精密挿入技能を得られることを示す。プロジェクトサイト https://mzhsoul.github.io/InsertAnything/ で、シミュレーションと実機実験のスクリプト、アセット、訓練済みチェックポイントを公開している。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Contact-rich precision insertion is a key manipulation skill in robotic assembly. Tight clearances make insertion more sensitive to alignment errors and prone to collisions and jamming, while variations in geometry and clearance across parts further complicate policy reuse. We present a reinforcement learning framework that trains insertion policies entirely in simulation for direct deployment without real-world demonstrations or policy fine-tuning. By combining target poses with compact three-dimensional fingertip force feedback, the policy learns to search for alignment and correct its motion despite errors in the estimated hole position. A decoupled gated reward coordinates alignment and insertion. Force-signal smoothing and state-independent standard deviations stabilize the learning process. The resulting policies perform real-world insertion across multiple hole geometries with a minimum nominal clearance of 0.02 mm and improve success while reducing peak contact forces under hole-position errors. Cross-clearance and cross-geometry evaluations further confirm policy generalization. The system achieved the first perfect score of 20/20 on ManipulationNet's peg-in-hole benchmark under its Human-in-the-Loop protocol, with fully autonomous insertion motions. A single policy trained only on a simulated hexagonal insertion task achieved an overall success rate of 95.0% across eight unseen real-world insertion tasks. These results show that learning entirely in simulation can yield precision insertion skills that can be deployed directly and reused across real-world tasks. The project website (https://mzhsoul.github.io/InsertAnything/) provides open-source simulation and real-robot experiment scripts, assets, and trained checkpoints.

arXiv ID: 2609.24511 / 要約の誤りについて