arXiv論文メモ
新着一覧
cs.RO / cs.AI / cs.LG · 査読状況未確認

ロボットの行動列の長さを予測の確かさで調整

GeoAAC: Geometry-Based Adaptive Action Chunking from Denoising Trajectories in VLA Policies

Xin Chen, Sen Chen, Yujuan Ding, Jian Liu, Guoqing Wang, Wei Ye, Heng Tao Shen, Yi Bin

この論文をやさしく読む

ひとことで言うと

ロボットが一度に実行する動作の長さを、予測の信頼性に合わせて変える手法です。フロー型VLA模型のノイズ除去の軌跡を使います。

何に役立つ?

連続して動く効率と、こまめに観測して修正する精密さのバランスを取る方法です。追加学習なしで一回の生成から動作区間を選びます。

この研究の面白いところ

軌跡の幾何的な変化が予測の不確かさと相関することを利用します。実機の平均成功率は比較条件で53.3%から74.4%へ改善しています。

どこまで分かった?

結果は指定された二つのVLAとシミュレーション・実機課題での検証です。幾何的な指標がどの課題でも信頼性を完全に表すという保証はありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

視覚・言語・行動(VLA)ポリシーでは、行動の生成と実行に行動チャンク化が広く用いられるが、既存手法は通常、行動ホライズンを固定している。一連の実行の途中では、タスクの段階ごとに、行動の連続性、制御精度、閉ループフィードバックに求められる水準が異なり得る。そのため、固定ホライズンでは変化する制御要件に対応できない。 本研究では、フローに基づくVLAポリシー向けに、現在の行動予測の信頼性に応じて行動ホライズンを調整する、幾何に基づく適応的行動チャンク化手法GeoAACを提案する。Flow Matchingのノイズ除去軌道の幾何は、予測の信頼性を特徴づける処理過程の情報を与え、行動列の先頭部分ごとの幾何的変動は予測の不確実性と正の相関を保つことを示す。GeoAACは、この先頭部分ごとの幾何を使ってホライズンごとの幾何的プロファイルを構築し、追加訓練なしに1回の生成から行動ホライズンを適応的に決める。 GR00T N1.5とπ0.5を用い、LIBERO、LIBERO-Pro、RoboCasa365、および実世界の物体操作タスクで実験した。その結果、固定ホライズンのベースラインと既存の適応手法に対して一貫した改善が得られた。改善幅はシミュレーションで最大8.7パーセントポイントであり、実世界での平均成功率は53.3%から74.4%へ上昇した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Action chunking is widely used for action generation and execution in Vision-Language-Action (VLA) policies, yet existing approaches commonly use a fixed action horizon. During a rollout, different task stages may require different levels of action continuity, control precision, and closed-loop feedback, making a fixed horizon unable to accommodate changing control requirements. We propose \textbf{GeoAAC}, a geometry-based adaptive action chunking method for flow-based VLA policies that adjusts the action horizon according to the reliability of the current action prediction. We show that the geometry of Flow Matching denoising trajectories provides process-level information for characterizing prediction reliability, with geometric variation across action prefixes remaining positively correlated with predictive uncertainty. GeoAAC uses this prefix-wise geometry to construct a horizon-wise geometric profile and adaptively determine the action horizon from a single generation without additional training. Experiments with GR00T N1.5 and {\pi}0.5 on LIBERO, LIBERO-Pro, RoboCasa365, and real-world manipulation tasks show consistent improvements over fixed-action-horizon baselines and existing adaptive methods, including up to 8.7 percentage points in simulation and an increase in average real-world success rate from 53.3\% to 74.4\%.

著者のコメント

9 pages, 6 figures. Submitted to the IEEE International Conference on Robotics and Automation (ICRA) 2027

arXiv ID: 2609.20776 / 要約の誤りについて