arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

対象の奥行き推定と予測制御で水中の目標を追跡

Underwater Visual Target Tracking with Target-Specific Depth Estimation and Adaptive Model-Fusion Predictive Control

Yuheng Zhou, Haiyang Cheng, Yanqi Feng, Pangkit Fong, Mei Xuan Lee, Marcus Gee, Chongrong Fang and Jianping He

この論文をやさしく読む

ひとことで言うと

水中ロボットが左右のカメラから対象までの距離を安定して見積もり、動く対象を追う方法です。

何に役立つ?

考えられる用途は、水中で深度計測が不安定なときの対象追従です。

この研究の面白いところ

対象らしい画素の選択とフィルタリングに加え、停止モデルと等速モデルの重みを予測誤差から調整して制御します。

どこまで分かった?

シミュレーションと実機の両方で既存方式に対する改善を報告しています。要旨には具体的な誤差値や評価環境の範囲は示されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

画像に基づく水中目標追跡では、奥行き測定の信頼性不足と、目標の動きが未知であることが課題となる。本論文では、自律型水中航行体(AUV)のためのステレオ視覚サーボの枠組みを提案する。 知覚では、対象に特化した奥行き抽出とカルマンフィルタリングを通じて、ステレオ画像から安定した三次元の相対状態を導く。色、視差、時間的な手がかりから対象の奥行きマスクを構成し、信頼できる対象画素を選ぶ。その後、得られた奥行き測定と、検出した画像上の中心をそれぞれ別にフィルタリングする。 制御では、ヨー角の調整を並進制御から分離することで、計算負荷の高い多自由度の結合最適化を回避し、並進運動のモデル予測制御(MPC)を実時間で実行できるようにする。並進制御器は適応的モデル融合予測制御を用い、等速度と速度ゼロの目標モデルを組み合わせて、異なる目標運動パターンに対応する。過去の予測誤差を使ってモデルの重みを更新し、駆動、追従距離、視野の制約の下で並進指令を計算する。シミュレーションと実環境での実験を通じて、提案する枠組みの有効性を検証し、既存の枠組みより優れた性能を示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Vision-based underwater target tracking is challenged by unreliable depth measurements and unknown target motion. This paper proposes a stereo visual-servoing framework for an autonomous underwater vehicle (AUV). For perception, the framework derives a stable 3D relative state from stereo images through target-specific depth extraction and Kalman filtering. It constructs a target-depth mask from color, disparity, and temporal cues to select reliable target pixels, and then filters the resulting depth measurement and detected image center separately. For control, the framework decouples yaw regulation from translational control, avoiding computationally expensive coupled multi-DOF optimization and enabling real-time translational MPC. The translational controller employs adaptive model-fusion predictive control, combining constant-velocity and zero-velocity target models to accommodate different target-motion patterns. It updates the model weights using historical prediction errors and computes translational commands subject to actuation, following-distance, and field-of-view constraints. Through simulations and real-world experiments, we validate the effectiveness of the proposed framework and show it has better performance than existing frameworks.

著者のコメント

9 pages,8 figures

arXiv ID: 2609.20731 / 要約の誤りについて