arXiv論文メモ
新着一覧
cs.RO / cs.AI / cs.LG / cs.SY / eess.SY · 査読状況未確認

異なるロボット間で共有する安全フィルターを潜在空間で学習

CrossSafe: Towards Cross-Embodiment Latent Safety Filters

Ihab Tabbara, Yuxuan Yang, Hussein Sibai

この論文をやさしく読む

ひとことで言うと

複数の双腕ロボットで安全判断を共有しながら、機体ごとの動き方の違いも考慮する方法です。

何に役立つ?

共通の操作方策を別形態のロボットへ移す際の衝突回避に役立つ可能性があります。要旨では五種類の双腕ロボットと五課題で衝突率の低下を報告しています。

この研究の面白いところ

安全な抽象行動の判断を共有しつつ、Hamilton–Jacobi到達可能性を形態を考慮した潜在空間で計算します。学習から除いた形態への適用も試しています。

どこまで分かった?

評価対象は五種類の双腕ロボットと五つの操作課題です。要旨には衝突率の具体的な数値や、ほかの種類のロボットでの結果はありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

異なる形態のロボットを横断する学習では、視覚・言語・行動モデルなど一つのモデルが、複数のロボットと課題に使える状態表現や操作技能を学べることが示されている。本研究は安全性の確保にも同じ考え方が通用すると仮定する。障害物を見つけ、避ける必要を認識し、安全な抽象的行動を選ぶ推論はロボット間でおおむね共通する。一方、形態、運動学、動力学によって、その行動を実現する方法と、安全で実行可能な動作は異なる。同じ行動が一方のロボットでは安全でも、別のロボットでは危険になり得る。これは特に、ロボットの形態や運動学に応じた安全性を明示的に表さず、共通のエンドエフェクター行動空間で動く汎用操作方策で重要となる。 提案するのは、ロボットの形態を条件とする安全フィルタリングである。Hamilton–Jacobi到達可能性に基づく価値関数と、安全性を最大化する対応方策をロボット間で共有する。ロボットと環境の形態を考慮した潜在表現を使い、潜在空間で直接Hamilton–Jacobi到達可能性解析を行う。これにより、安全性の概念を異なる形態へ一般化しつつ、各ロボットの形態と運動学を明示的な条件として残す。全身の衝突回避制約がある五種類の双腕ロボットと五つの操作課題で評価した。五課題と四種類の形態で共同学習した一つの方策は、学習から除いた形態へ追加学習なしで適用でき、元の方策の衝突率を下げた。学習に使う形態を増やすと一般化も改善した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Cross-embodiment learning has shown that a single model, such as a vision-language-action (VLA) model, can learn state representations and manipulation skills that can be applied across heterogeneous robots to accomplish various tasks. We hypothesize that the same holds for safety enforcement. The reasoning required to satisfy a safety constraint, such as detecting an obstacle, recognizing that it should be avoided, and selecting a safe abstract action, is largely shared across robots. What differs across embodiments is how the abstract safe action is realized: morphology, kinematics, and dynamics determine which actions are safe and feasible. Consequently, the same action can be safe for one robot and unsafe for another. This is especially important for generalist manipulation policies that operate in a common end-effector action space without explicitly capturing how safety depends on the robot's morphology and kinematics. We propose embodiment-conditioned safety filtering, in which a Hamilton-Jacobi reachability-based value function and its corresponding safety-maximizing policy are shared across robots. Using a morphology-aware latent representation of the robot and its environment, we perform Hamilton-Jacobi reachability analysis directly in latent space so that the learned safety concepts can generalize across embodiments while remaining explicitly conditioned on each robot's morphology and kinematics. We evaluate our approach across five bimanual robot embodiments and five manipulation tasks with whole-body collision-avoidance constraints. Our results show that a single policy, jointly trained across five manipulation tasks and four embodiments, exhibits zero-shot generalization to a held-out embodiment, reducing the nominal policy's collision rate. They also show that training using more embodiments improves generalization.

arXiv ID: 2609.28984 / 要約の誤りについて