arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

受動車輪付き飛行ロボットの空陸切替と路面追従を学習

Learning Air-Ground Motion Control with Temporal Mode Switching and Cross-Terrain Tracking

Ruitian Pang, Mingrui Li, Xuanting Liu, Tiancheng Lai, Juncheng Chen, Xiangyu Li, Ruibin Zhang, Qishao Wang, Jin Yu, Haiyin Piao, Fei Gao, Chao Xu, Yanjun Cao

この論文をやさしく読む

ひとことで言うと

飛行と車輪での地上移動を切り替えるロボットで、少ない距離情報でも移動モードを選び、路面が変わっても軌道に沿って進む制御を学習しています。

何に役立つ?

考えられる用途は、地上で省エネルギーに移動し、必要な区間を飛行する空陸ロボットの制御です。シミュレーションだけでなく実環境でも評価しています。

この研究の面白いところ

単一点の距離測定でも、その履歴とこれから進む軌道の情報を使って切替を判断します。モード選択と軌道追従を組み合わせた101 mの走行・飛行も報告しています。

どこまで分かった?

位置RMSE 0.08 mは報告された101 m軌道での結果です。PIDより低誤差という比較は試した条件の範囲であり、すべての路面や障害物に対応するとまでは示していません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

受動車輪を備える地上・空中二形態車両(TABV)は、空中での移動能力と省エネルギーな地上移動を組み合わせる。しかし実環境への応用では、搭載知覚が限られた状況での確実な空陸モード切替と、多様な路面にわたる頑健な地上軌道追従が依然として課題である。本研究では、受動車輪付きTABVのための学習ベースの空陸移動制御枠組みを提案する。 (1)空陸の移動モードを自律的に切り替える学習済みモード選択器。過去の単一点飛行時間方式(ToF)の測距値とロボット状態を、将来の参照情報と合わせて用い、使用する移動モードを決める。(2)軌道追従のための強化学習制御方策。自己状態に関する観測と将来の参照情報を組み合わせ、軌道の変化を先取りする。地上移動では、複数路面での訓練と動力学のランダム化により、異なる路面で頑健な追従を可能にする。 シミュレーションと実環境実験により、限られた知覚での確実な空陸切替と、多様な路面条件での正確な地上追従を示す。学習した選択器は難しい遷移で規則ベースの選択器を上回る。地上制御器は、試したすべての条件でPIDより位置の二乗平均平方根誤差(RMSE)が小さく、非線形モデル予測制御(NMPC)が失敗する場面でも良好な追従を保つ。これらの能力を統合したシステムは、複数回の自律的モード遷移を含む101 mの空陸軌道を、位置RMSE 0.08 mで追従する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Passive-wheeled terrestrial-aerial bimodal vehicles (TABVs) combine aerial mobility with energy-efficient ground locomotion. However, reliable air-ground mode switching under limited onboard perception and robust ground trajectory tracking across diverse terrains remain challenging when targeting real-world applications. In this work, we propose a learning-based air-ground motion control framework for passive-wheeled TABVs: 1) a learned mode selector for autonomous air-ground motion mode switching. The selector uses historical single-point time-of-flight (ToF) measurements and robot states together with future reference information to determine the active locomotion mode. 2) a reinforcement learning control policy for trajectory tracking. The policy combines proprioceptive observations with future reference information to anticipate trajectory changes. For ground locomotion, multi-terrain training and dynamics randomization enable robust tracking across different terrains. Simulation and real-world experiments demonstrate reliable air-ground switching under limited perception and accurate ground tracking across diverse terrain conditions. The learned selector outperforms a rule-based mode selector in challenging transitions, while the ground controller achieves lower position RMSE than PID across all tested conditions and maintains decent tracking where NMPC fails. With these capabilities integrated, the system tracks a 101m air-ground trajectory through multiple autonomous mode transitions with a position RMSE of 0.08m.

arXiv ID: 2609.26564 / 要約の誤りについて