arXiv論文メモ
新着一覧
cs.CV / cs.SY / eess.SY · 査読状況未確認

ロボットの動作結果を予測して補正を選ぶ方法

CereVLA: Cerebellum-Inspired Consequence-Aware Residual Governance for Efficient Vision-Language-Action Execution

Shuai Zeng, Yuxuan Liang, Hangmiao Hu, Fobao Zhou, Zixiang Wang, Wenxi Hong, and Hang Zhao

この論文をやさしく読む

ひとことで言うと

ロボットの動作のずれを直す際、補正後に何が起きそうかを予測し、悪い結果につながりそうな補正を抑える方法。

何に役立つ?

既存のVLA方策を再訓練せず、連続動作の実行中に補正を加える仕組みの設計に役立つ可能性がある。報告された成功率の改善は、要旨にある評価条件での結果である。

この研究の面白いところ

補正を作る処理と、その短期・区間先の結果を評価する処理を分け、予測に基づいて補正そのものを止められる。SO-101では成功率と成功試行のステップ数の両方を比較している。

どこまで分かった?

要旨に記された具体的な数値はSO-101で凍結したSmolVLAとの比較である。LIBEROでの改善幅や、別の作業環境での性能は要旨には示されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

複数の動作をひとまとまりで出力する視覚・言語・動作(VLA)方策は推論を効率化できるが、実行を確定した動作列の途中で十分なフィードバックを得られず、誤差が蓄積することがある。残差による適応はVLAを再訓練せずにずれを直せるものの、従来の補正は基準動作との一致を重視し、その後の結果を明示的には考慮していない。 著者らは、凍結したVLAの実行に、軽量な残差の精緻化と予測される結果の評価を組み合わせる統一的な枠組みCereVLAを提案する。まずフローに基づく残差の精緻化で補正動作を生成し、次に再帰的な状態空間モデルと履歴を考慮する分類器で、短期および一定区間先の結果を評価する。望ましくない結果が予測された補正は、軽量な制御部が選択的に抑える。LIBERO-10とLIBERO-GOALでは先行する高性能手法と比較し、CereVLAの有効性を示した。SO-101では、凍結したSmolVLAを基準に、課題成功率を57.5%から90.0%に高め、成功試行における平均制御ステップ数を19.6%減らした。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Action-chunked vision-language-action (VLA) policies improve inference efficiency, but limited feedback within committed action chunks can lead to accumulated execution errors. Residual adaptation can correct such deviations without retraining the VLA; however, existing corrections are typically optimized for reference-action consistency without explicitly considering their downstream consequences. To address this limitation, we present Cerebellum-Inspired Consequence-Aware Residual Governance (CereVLA), a unified framework that integrates lightweight residual refinement and predictive consequence evaluation into frozen VLA execution. Corrective actions are first generated by flow-based residual refinement, and their short- and interval-horizon consequences are then evaluated by a recurrent state-space model and a history-aware classifier. Residual corrections predicted to be unfavorable are selectively suppressed by a lightweight governor. Comparisons with state-of-the-art methods on LIBERO-10 and LIBERO-GOAL demonstrate the effectiveness of CereVLA. On SO-101, CereVLA increases task success from 57.5% to 90.0% and reduces mean control steps by 19.6% among successful trials, relative to the frozen SmolVLA baseline.

著者のコメント

8 pages, 5 figures

arXiv ID: 2609.27468 / 要約の誤りについて