追加学習なしで水中を移動・操作するロボット制御
AquaCap: A Training-Free Underwater Embodied Agent with Code-as-Policy
この論文をやさしく読む
ひとことで言うと
追加学習なしで、水中ロボットの移動と物体操作を計画・修正する仕組み。
何に役立つ?
水中で学習データを大量に集めにくい作業の自律化を検討する参考になる。
この研究の面白いところ
行動の失敗を記憶して制御コードを修正し、位置がずれた対象にも対応した点。
どこまで分かった?
シミュレーションの成功率は66.43%。実機で把持と搬送を示したが、実機の成功率は要旨にない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
視覚・言語・行動モデルの進歩により水中の身体を持つ知能への関心が高まっているが、大量の相互作用データへの依存は、データ収集が高価で数も限られる水中での適用を妨げる。本研究は、追加学習を必要としないCode-as-Policyの枠組みAquaCapを提案し、自律的な水中移動と物体操作を扱う。二層のエージェントが作業指示と環境観測を、条件を考慮した計画と実行可能な制御プログラムへ変換する。構造化された知覚は、水中で観測品質が悪化した条件でも、意味、幾何、信頼性を考慮した観測をエージェントに与える。失敗を考慮する記憶は、成功しなかった行動を診断し、閉ループでの再計画とコード修正を支える。これにより作業別の学習やパラメータ更新なしにオンラインで適応する。シミュレーションでの成功率は66.43%だった。実環境の実験では、遠隔操作水中機を使った自律的な把持と物体搬送を示し、水流による外乱で位置が変わった対象も操作した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-19(UTC)
- 最新改訂
- 2026-09-19 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-19 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Recent advances in vision-language-action models have stimulated growing interest in underwater embodied intelligence. However, their reliance on large-scale interaction data limits their applicability underwater, where data collection is costly and scarce. To address this challenge, we present AquaCap, a training-free Code-as-Policy framework for autonomous underwater navigation and manipulation. AquaCap employs a dual-layer agent that translates task instructions and environmental observations into condition-aware plans and executable control programs. Structured perception then provides the agent with semantic, geometric, and reliability-aware observations under degraded underwater conditions. A failure-aware memory diagnoses unsuccessful actions and supports closed-loop replanning and code revision. This design enables online adaptation without task-specific training or parameter updates. AquaCap achieves a 66.43% success rate in simulation. Real-world experiments further demonstrate autonomous grasping and object transport with an ROV, including the manipulation of targets displaced by hydrodynamic disturbances.
著者のコメント
8 pages, 6 figures, 4 tables, 1 algorithm
arXiv ID: 2609.23133 / 要約の誤りについて