言語モデルの操作計画と最適制御を組み合わせた狭所駐車
From Semantic Decisions to Feasible Trajectories: Self-Evolving LLM-Guided Optimal Control for Narrow-Space Parking
この論文をやさしく読む
ひとことで言うと
言語モデルには前進や切り返しなどの大まかな判断を任せ、実際の軌道は車両の制約を扱う最適化で求める駐車手法です。失敗した理由を次の計画や知識ベースに反映します。
何に役立つ?
考えられる用途は、狭い場所で最適化が失敗した際の操作順序の見直しです。言語的な判断を、実行可能な軌道を求める計算へ接続する方法を示しています。
この研究の面白いところ
オンラインの再計画だけでなく、失敗経験をオフラインで構造化した知識に変える仕組みがあります。同じ操作の表現を、自動車型と差動駆動という異なる運動学で検証しています。
どこまで分かった?
要旨は自動車型モデルのシミュレーションと差動駆動ロボットによる検証を述べますが、成功率、比較対象、試行数は示していません。一般道路での自動車の安全性が確立したとは判断できません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
非凸で狭い環境での自動駐車は、依然として難しい。最適制御手法は、車両の運動力学と衝突制約を明示的に課せるが、非凸性がソルバーの頑健性を損ない、失敗につながることがある。大規模言語モデル(LLM)は高い意味的推論能力を示す一方、密な軌道を直接生成させると、物理的に実行可能であることを保証しにくい。 本研究では、LLMが上位の離散的な操作判断を行い、最適制御モジュールが下位の車両運動力学と衝突制約を満たす、統合的な枠組みSE-LLM-OCPを導入する。オンラインでは、LLMが疎な操作計画を提案し、駐車課題を短い予測区間の軌道最適化問題の列に分解する。次に、下位のソルバーが最適制御問題を順番に解く。ソルバーが失敗した場合、LLMはソルバーと検証段階から失敗の証拠を集約し、再計画に利用する。オフラインでは、蓄積されたオンラインの失敗を基に、SE-LLM-OCPが構造化された意思決定知識ベースをゼロから自動的に発展させる。 提案した枠組みを、自動車型の車両モデルによるシミュレーションと、差動駆動ロボットで検証する。実験結果は、SE-LLM-OCPが狭い状況でより安全な自動駐車を可能にし、同じ操作表現を運動学の異なるプラットフォームへ移せることを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-21(UTC)
- 最新改訂
- 2026-09-21 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-21 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Autonomous parking in nonconvex and narrow environments remains challenging. Although optimal-control methods can explicitly enforce vehicle dynamics and collision constraints, nonconvexity compromises solver robustness and can cause failures. Large language models (LLMs) exhibit strong semantic reasoning capabilities, but directly generating dense trajectories makes it difficult to guarantee physical feasibility. We introduce SE-LLM-OCP, a unified framework in which LLMs make high-level discrete maneuver decisions, while an optimal-control module enforces low-level vehicle dynamics and collision constraints. Online, the LLM proposes sparse maneuver plans, decomposing the parking task into a sequence of short-horizon trajectory-optimization problems. A low-level solver then sequentially solves optimal-control problems. If the solver fails, the LLM aggregates failure evidence from the solver and validation stages to guide replanning. Offline, SE-LLM-OCP automatically evolves a structured decision-making knowledge base from scratch, driven by accumulated online failures. We validate our proposed framework in simulation on a car-like vehicle model and on a differential-drive robot. Our experimental results show that SE-LLM-OCP enables safer autonomous parking in narrow scenarios and demonstrates transfer of the same maneuver representation to a different kinematic platform.
arXiv ID: 2609.24631 / 要約の誤りについて