地形変化に対応して連続掘削する自律油圧ショベル
From Target Selection to Digging: A Learning-Based Framework for Continuous Autonomous Excavation
この論文をやさしく読む
ひとことで言うと
形が変わる土の山を見ながら、次に掘る場所とショベルの動きを調整する仕組み。
何に役立つ?
連続掘削を自動化する制御方法の設計や、目標選択と掘削動作の役割分担を検討する材料になる。
この研究の面白いところ
LiDARで掘る場所を選び、強化学習と模倣学習を組み合わせ、縮尺モデルの実機で積載量を比較した。
どこまで分かった?
実証は縮尺モデルの油圧ショベルと記載された試行条件で行われた。実際の大型建機や長期運用での性能は要旨に示されていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
掘削を繰り返すと土の山の形が常に変わるため、自律ショベルは掘削対象を適応的に選び、連続した掘削サイクルを通して動作を調整する必要がある。著者らは、地形を考慮した目標選択と強化学習・模倣学習の制御器を統合した、連続自律掘削の学習ベースの枠組みを提示する。この枠組みでは、目標に応じた移動と局所的な掘削を分ける。共通のタスク条件付き強化学習方策が経由点に沿った接近と荷を積んだ状態での運搬を制御し、模倣学習方策は専門家の実演から視覚に基づく掘削と持ち上げを学ぶ。掘削目標はLiDARの標高マップから選び、動作制御用のバケット先端の経由点に変換する。制御構成は共通の動作インターフェースを介して学習済み方策と決定論的な荷下ろしを協調させる。全体のシステムを、複数種類のセンサーと閉ループのアクチュエーター制御を備えた縮尺モデルの油圧ショベルに実装した。オフラインの再生評価と実機実験では、各比較手法に比べて目標選択がより一貫し、局所的な移動時間が短く、積載量が増えた。学習した掘削方策の完了サイクル当たり平均積載量は6.52 kgで、固定掘削方式の2.68 kgと比べて高かった。さらに、各5回すくう試行を3回行い、土の山の形が変わり続ける中で連続自律掘削ができることを示した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Repeated excavation continuously reshapes pile geometry, requiring an autonomous excavator to adapt its digging targets and coordinate motion across successive excavation cycles. We present a learning-based framework for continuous autonomous excavation that integrates terrain-aware target selection with reinforcement- and imitation-learning controllers. The framework separates target-conditioned motion from local digging: a shared task-conditioned RL policy controls waypoint-guided approach and loaded transport, while an IL policy learns vision-based digging and lifting from expert demonstrations. Digging targets are selected from LiDAR elevation maps and converted into bucket-tip waypoints for motion control. The control architecture coordinates the learned policies and deterministic unloading through a shared motion interface. The complete system is deployed on a scaled hydraulic excavator with multimodal sensing and closed-loop actuator control. Offline replay and physical experiments demonstrate more consistent target selection, shorter local motion time, and increased payload compared with the respective baselines. The learned digging policy achieves a mean payload of 6.52 kg per completed cycle, compared with 2.68 kg for Fixed Dig. Three five-scoop runs further demonstrate consecutive autonomous excavation under continuously changing pile geometry.
著者のコメント
8 pages, 7 figures, 4 tables
arXiv ID: 2609.29750 / 要約の誤りについて