実行記録を検証して改善するロボットエージェント基盤RegenHarness
RegenHarness: A Robot Agent Harness with Evidence-Gated Recursive Self-Improvement
この論文をやさしく読む
ひとことで言うと
ロボットの提案・実行・完了確認を分け、証拠に基づいて作業状態を更新する基盤である。
何に役立つ?
長時間のロボット作業で、誤った完了報告や設定変更を防ぐ設計に役立つ可能性がある。
この研究の面白いところ
作業記録から改善案を作るが、回帰検査と承認を通し、モデルの重みや完了確定の条件は勝手に変えない。
どこまで分かった?
四足歩行ロボットの倉庫作業などの事例を示した。広範な作業での成功率や一般化性能の数値は要旨にない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
長時間にわたるロボットの作業では、モデルが提案したこと、制御器が終了したこと、作業が検証済みで完了したことを明確に分ける必要がある。本研究は、作業計画をさまざまなロボット技能につなぐ、証拠で進行を制限するロボットエージェント基盤RegenHarnessを提示する。実行構成では、文脈に応じた提案を行うモデルのループと、指示の配送、観測、検証、確定、範囲を限った復旧を行うエージェントのループを結び付ける。役割を分離した四つの文脈で、計画、監督、検証、復旧の入力を区別する。版を付けた記憶は観測事実と承認済みの進捗を区別し、識別情報と版に結び付いた確定ゲートが信頼された作業状態への更新を制御する。実行環境は、重複した指示の抑制、資源の使用権、復旧予算を明示的なバックエンド契約のもとで組み合わせ、完了を報告する前に元の利用者の目標を確認する。 著者らの知る限り、身体を持つロボットエージェントに対し、証拠を条件とする再帰的自己改善(RSI)の手順を導入した最初の研究である。複数の作業の実行記録から、文脈規則、作業テンプレート、配送、復旧方針の変更案を作る。固定された回帰検査とリリース承認で採否を決め、版付きの展開と巻き戻しで設定の追跡可能性を保つ。このRSI手順は基盤の設定を改訂するが、オンラインでモデルの重みを更新したり、確定ゲートを弱めたりはしない。実際の四足歩行ロボットへの導入例では、音声で開始する倉庫内移動、全方位の点検、視覚解析、メッセージの配達、帰還、音声での報告を、音声、画像、軌跡、受領記録で結び付けて記録する。別の一連の作業では、目的地点への近さだけでなく実行履歴によって完了が決まる理由を示す。これらの事例は、現実のロボット作業で知覚、身体動作、コミュニケーション、履歴に依存した完了判定を統合したことを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-23(UTC)
- 最新改訂
- 2026-09-23 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-23 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Long-horizon robot execution requires a clear distinction between a model's proposal, a controller's termination, and verified task completion. We present RegenHarness, an evidence-gated robot-agent harness connecting task planning to heterogeneous robot skills. Its execution architecture couples a model loop for context-conditioned proposals with an agent loop for dispatch, observation, verification, commitment, and bounded recovery. Four role-isolated contexts separate planning, supervision, verification, and recovery inputs. Versioned memory distinguishes observed facts from accepted task progress, while an identity- and version-bound commit gate controls updates to trusted task state. The runtime combines duplicate-dispatch control, resource leases, and recovery budgets under explicit backend contracts, and checks the original user goal before reporting completion. To our knowledge, we are the first to introduce an evidence-gated recursive self-improvement (RSI) protocol for embodied robotic agents. Across missions, execution records motivate candidate changes to context rules, task templates, routing, and recovery policies; fixed regression checks and release authorization govern their acceptance; versioned rollout and rollback preserve configuration traceability. This RSI protocol revises the harness configuration without online model-weight updates or permission to weaken the commit gate. A real quadruped deployment documents voice-triggered warehouse navigation, panoramic inspection, visual analysis, message delivery, return, and spoken reporting through linked audio, images, trajectories, and receipts. A separate circuit demonstrates why completion depends on execution history rather than endpoint proximity alone. Together, the cases demonstrate integrated perception, physical execution, communication, and history-dependent completion in real-world robot tasks.
arXiv ID: 2609.27612 / 要約の誤りについて