arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

水中ロボットが標的を見失っても周回を続ける制御

AquaOrbit: Sim-to-Real Reinforcement Learning for Underwater Target Orbiting under Intermittent Visual Feedback

Kanzhong Yao, Jinyi Leng, Hao Zhang, Zhe Sun and Xuelong Li

この論文をやさしく読む

ひとことで言うと

水中ロボットが標的を一時的に見失っても、直前の方向や深度を利用して安定を保ち、再び標的を見つける制御です。

何に役立つ?

考えられる用途は、水中対象物の周囲を回りながら観察するロボットです。実機では機上処理だけで複数の軌道を実演しています。

この研究の面白いところ

別のシミュレータへの移行と実機への移行を分けて検証し、復旧モジュールを外した比較もあります。視覚がある間の追従だけでなく、失った後の立て直しを組み込んでいます。

どこまで分かった?

20回中20回の完了と約46%の誤差低減はシミュレータ間評価の結果です。実機の8秒遮蔽と2.5秒以内の再捕捉は報告された条件での結果で、あらゆる見失いへの保証ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

水中で標的の周囲を回る際、視覚が断続的に失われると標的に対するフィードバックが途切れ、協調した運動の維持と動く標的の再捕捉が難しくなる。本研究では、視覚フィードバックが中断する水中標的周回のために、復旧モジュールを備えた強化学習制御器AquaOrbitを提示する。検出が失われている間、復旧モジュールは保持しておいた視線方向、ロール角、深度の参照値を使い、姿勢などの安定化と標的の再捕捉を支える。制御器は、動力学、観測、視覚喪失をランダム化したIsaac Simで学習する。 異なる物理エンジンと知覚の摂動を持つGazebo/ROS2で、再学習なしに評価したところ、未学習の深度可変3次元軌道で、静止標的と移動標的の各条件につき20回中20回の周回試行を完了した。移動標的条件では、復旧機能を持つPIDベースの視覚サーボ制御器に対し、同等の経路追従精度を保ちながら平均視線誤差を約46%低減した。復旧モジュールを除くと完了は20回中9回に低下した。知覚と制御を完全に機上で行う実機へのゼロショット展開では、楕円、8の字、深度可変の円軌道を実演した。後二者の軌道種別は学習に含まれていない。ロボットは手動で最大8秒間遮蔽した間も姿勢を安定に保ち、報告された、姿勢変化により視野から標的が外れた事象では、2.5秒以内に標的を再捕捉した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Intermittent visual loss disrupts target-relative feedback during underwater orbiting, making it difficult to maintain coordinated motion and reacquire a moving target. We present AquaOrbit, a reinforcement-learning controller with a recovery module for underwater target orbiting under interrupted visual feedback. During detection loss, the recovery module uses latched line-of-sight, roll, and depth references to support stabilization and target reacquisition. We train the controller in Isaac Sim with dynamics, observation, and vision-loss randomization. Evaluated without retraining in Gazebo/ROS2 under a different physics engine and perception perturbations, AquaOrbit completes 20/20 orbiting trials in each of the static- and moving-target conditions on an unseen variable-depth 3-D trajectory. In the moving-target condition, it reduces mean line-of-sight error by approximately 46% relative to a PID-based visual servoing controller with recovery while maintaining comparable path-tracking accuracy; removing the recovery module reduces completion to 9/20. Zero-shot physical deployment with fully onboard perception and control demonstrates elliptical, figure-eight, and variable-depth circular trajectories, including the latter two trajectory types absent from training. The robot maintains attitude stability during manual occlusions lasting up to 8s and reacquires the target within 2.5s in the reported attitude-induced field-of-view loss events.

著者のコメント

This work has been submitted to IEEE for possible publication

arXiv ID: 2609.24054 / 要約の誤りについて