記憶と経験の更新で動作を改善する身体性エージェント
ME-Brain-1.0: Memory, Cognition and Action for Evolving Embodied Intelligence
この論文をやさしく読む
ひとことで言うと
動作の経験を記憶にまとめ、次の判断や実行に利用して、モデルの再学習なしに改善を図るシステムです。
何に役立つ?
一度学習した能力を固定するだけでなく、蓄積した経験を再利用する身体性エージェントの設計に役立ちます。
この研究の面白いところ
経験の保存、スキルへの変換、行動生成を分け、判断に重要な時点や領域へ計算を集中させます。モデルの重みを再学習することとは異なる更新を扱います。
どこまで分かった?
評価は挙げられたベンチマークで、ME-RealBenchは6タスクです。スコアと成功率は別指標で、報告された差も相対的な増加率ではなくポイント差です。無制限の長期自己改善を実証したものではありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
現在の身体性システムは主として、導入後も固定された事前学習済み能力に依存し、物理的な相互作用から学ぶ能力が制約されている。本研究では、行動実行、経験獲得、経験の進化、改善された実行という閉ループを中心に構成する、自己進化型の身体性システムMachEmbodied-Brain(ME-Brain)を導入する。Evolvable Memoryは、マルチモーダルな軌跡を階層的で再利用可能な経験へ統合する。Cognitive Coreは、物理的な経験を転用可能なスキルへ変える。Action Modelは、イベント駆動のキーフレーム、EventCellによる局所世界の予測、行動条件付きの記憶変調を組み合わせ、判断上重要な時点、領域、過去の証拠に計算を集中させる。これらのモジュールによって、モデルの再学習なしで、学習して固定する方式から、導入後に進化する方式へ身体性知能を移行させる。 Cognitive Coreは、身体性およびエージェントのベンチマークで、最も強い比較モデルをそれぞれ8.2ポイントと9.6ポイント上回った。Action ModelはRoboMMEで平均成功率47.88%を達成し、最も強いベースラインより3.26ポイント改善した。RoboDojoでは平均スコア21.51、成功率16.03%を達成し、π₀.₅をそれぞれ10.10ポイントと9.12ポイント上回った。6タスクのME-RealBenchでは、ME-Brainは平均スコア69.5、成功率66.7%を達成し、DM0.5をそれぞれ12.8ポイントと11.7ポイント上回った。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-21(UTC)
- 最新改訂
- 2026-09-21 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-21 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Current embodied systems largely rely on pretrained capabilities that remain fixed after deployment, limiting their ability to learn from physical interaction. We introduce MachEmbodied-Brain (ME-Brain), a self-evolving embodied system organized around a closed loop of action execution, experience acquisition, experience evolution, and improved execution. Evolvable Memory consolidates multimodal trajectories into hierarchical, reusable experience; Cognitive Core transforms physical experience into transferable skills; and the Action Model combines event-driven keyframes, EventCell local-world prediction, and action-conditioned memory modulation to focus computation on decision-critical moments, regions, and historical evidence. Together, these modules shift embodied intelligence from train-and-freeze to deploy-and-evolve without model retraining. Cognitive Core outperforms the strongest comparison models by 8.2 and 9.6 points on embodied and agent benchmarks. The Action Model achieves 47.88% mean success on RoboMME, a 3.26-point improvement over the strongest baseline. On RoboDojo, it reaches a 21.51 mean Score and 16.03% success rate, exceeding $\pi_{0.5}$ by 10.10 and 9.12 points. On the six-task ME-RealBench, ME-Brain achieves a 69.5 mean Score and 66.7% success rate, outperforming DM0.5 by 12.8 and 11.7 points, respectively.
arXiv ID: 2609.24271 / 要約の誤りについて