arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

空と地上のロボットが観測能力に応じて探索を分担

HEROIC: Heterogeneous Evidential Reasoning for Open-Vocabulary Identification and Cross-Robot Collaboration

Mihir Chauhan, Aarav Jain, Addison Zucek, Manmeet Dang, Damon Conover, Aniket Bera

この論文をやさしく読む

ひとことで言うと

飛行ロボットと地上ロボットが、対象を見つけられる距離に応じて探索や案内の役割を切り替える方法です。自然言語で連携します。

何に役立つ?

捜索救助や危険環境の探索が想定用途です。空からの広域探索と地上での近距離確認を組み合わせる設計に役立ちますが、災害現場での成果そのものを示したという意味ではありません。

この研究の面白いところ

対象を識別できる高度が安全飛行高度を下回ると、飛行機は探索から地上機の案内などに転じます。発見したという判断にも近距離検証を要求します。

どこまで分かった?

6シーンの統合実験で到達率84%、比較手法35〜54%、到達の速さは2〜4倍と報告しています。要旨だけではシーン外への一般化や試行数の詳細は分かりません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

異種の空中・地上ロボットからなるマルチエージェントチームは、開かれた世界での探索に有望であり、偵察、都市捜索救助(USAR)、災害対応と復旧、危険環境などへの応用がある。この2種類のプラットフォームでは、うまく機能しない状況が異なる。空中ロボットは広範囲を素早く移動できるが、高所からは小さい対象や遮蔽された対象を識別できない。一方、地上ロボットは人や危険物などの対象を近距離で識別できるが、探索できる面積が小さい。既存の言語指示型チームは、役割が事前に固定されているか、人手で書かれた能力タグから言語モデルが役割を割り当てるため、任務中にどの時点である機体が役立たなくなったかを判断できない。 本研究では、自然言語だけでエージェント間の通信を行う、分散型の異種マルチエージェント・オープンボキャブラリ探索調整フレームワークHEROICを提案する。初期の役割割当ては、センサーの特性と、高い確信度で対象を検出できるかを判定するスケール則から導く。この法則は、任務の自然言語プロンプトだけから飛行高度と掃引間隔を割り当てる。計算された高度が安全飛行に必要な高度を下回る場合、空中エージェントは自ら役割を変更し、探索者から、地上エージェントのための上空からの優先順位付け、随伴、経路案内へ移る。両ロボットは、探索領域について証拠に基づく信念を保持する。正の証拠には方位を示す半直線を、負の証拠には対数オッズで表した事後分布を用い、到着の判定には必ず近距離での検証を要求する。 全構成要素を統合した実験では、6つの全シーンを通してHEROICは84%の割合で対象に到達した。同じ認識処理を用いる視覚言語フロンティア手法、フロンティア探索、往復掃引、ランダムウォークのベースラインは35〜54%であり、HEROICは対象への到達も2〜4倍速かった。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Multi-agent heterogeneous air-ground robot teams are attractive for open world search, with applications for reconnaissance, urban search and rescue missions (USAR), disaster response and recovery, and hazardous environments. These two platforms have different failure modes: aerial robots cover ground quickly but cannot resolve small or occluded targets from altitude, while ground robots can identify objects-of-interest, such as people or hazardous objects, at close range but cover less area. Existing language-tasked teams either have roles fixed prior, or have a language model assign them from hand-written capability tags, so the team is unable to know when within a mission an asset is no longer useful. We present HEROIC, a decentralized heterogeneous multi-agent open-vocabulary search coordination framework that requires agents to communicate in natural language only. HEROIC's initial agent role assignment is derived from sensor properties and a scale law to determine whether targets can be detected with a high confidence. From the mission's natural language prompt alone, this law assigns aerial flight altitudes and sweep spacing. When this calculated height falls below the altitude for safe flight, aerial agents re-task themselves from searcher to aerial triage, escort, and route guide for ground agents. Both robots maintain an evidential belief over the search area (bearing rays for positive evidence, a log-odds posterior for negative evidence) and gate any arrival on close-range verification. In full-stack experiments, HEROIC reaches the target 84% of the time across all 6 scenes, compares to 35-54% for vision-language frontier baselines, frontier-based search, lawnmower, and random-walk running the same perception, all while being 2-4x sooner to arrive at the target.

arXiv ID: 2609.19803 / 要約の誤りについて