手と物体の幾何を分けて扱う把持動作の生成法
DEAL-Grasp: Decoupled Alignment Representation for Geometry-Aware Dexterous Grasp Generation
この論文をやさしく読む
ひとことで言うと
手全体の位置と指の曲げ方を分けて表現し、物体をつかむ姿勢を生成する方法を提案した。
何に役立つ?
考えられる用途は、仮想現実やロボットの把持動作を、多様性と物理的な自然さを保ちながら生成すること。
この研究の面白いところ
推論時の追加最適化を使わず、学習したベクトル場の積分だけで把持を作り、二つのベンチマークで評価した。
どこまで分かった?
要旨ではMultiDexとゼロショットRealDexでの比較を述べるが、成功率や遅延の具体的な数値は記されていない。実機での評価も記載されていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
現実的な関節を持つ手と物体の相互作用を合成することは、仮想現実、身体性を持つ知能、デジタルヒューマンの応用における基本的な課題である。器用な把持を合成する従来法は通常、全体の剛体運動と局所的な関節運動を結びつけた関節空間で姿勢を回帰またはノイズ除去するため、生成例が不安定になり、接触が物理的に不自然になることが多い。本研究は、分離した位置合わせ(DEAL)表現に基づくDEAL-Graspを導入し、把持合成を位置合わせ空間での生成として捉え直す。相互作用の状態を、作業空間内の幾何学的な基準点と関節パラメータで構成し、局所的な関節の形を保ったまま、閉形式のプロクラステス位置合わせから剛体変換を復元する。 この混合状態で、成分ごとのベクトル場を持つ異種状態のフローマッチングを用いて把持の生成をモデル化し、学習中には時間に応じて変わる物理的な正則化を取り入れる。推論時は、学習済みのベクトル場を積分するだけで把持を合成し、推論時の最適化や補助的な物理ガイダンスは使わない。MultiDexとゼロショットのRealDexベンチマークでは、力による摂動に対する成功率が高く、物体へのめり込みが小さく、多様な把持を生成した。また、最適化への依存が大きい基準手法より、本来の推論遅延を大幅に短縮した。プロジェクトページは https://wmtlab.github.io/DEAL-Grasp/ で公開されている。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-23(UTC)
- 最新改訂
- 2026-09-23 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-23 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Synthesizing realistic articulated hand-object interactions is a fundamental problem in virtual reality, embodied intelligence, and digital human applications. Existing methods for dexterous grasp synthesis typically regress or denoise poses in a joint space that couples global rigid motion with local articulation, which often yields unstable samples and physically implausible contacts. We introduce DEAL-Grasp, built upon the Decoupled Alignment (DEAL) representation, which reformulates grasp synthesis as alignment-space generation: the interaction state comprises task-space geometric anchors and articulation parameters, from which the rigid transform is recovered via closed-form Procrustes alignment while preserving local articulation. On this mixed state, we model grasp generation using heterogeneous-state flow matching with component-wise vector fields, incorporating time-adaptive physical regularization during training. At inference, grasps are synthesized solely by integrating the learned vector field, without test-time optimization or auxiliary physical guidance. Across MultiDex and zero-shot RealDex benchmarks, DEAL-Grasp attains high force-perturbation success rates alongside minimal penetration and high diversity of generated grasps, while substantially reducing native inference latency compared to optimization-heavy baselines. The project page is available at https://wmtlab.github.io/DEAL-Grasp/.
arXiv ID: 2609.28131 / 要約の誤りについて