arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

記憶と経験の更新で動作を改善する身体性エージェント

ME-Brain-1.0: Memory, Cognition and Action for Evolving Embodied Intelligence

Wei He, Hengtao Li, Zhongrui Yu, Xuhan Zhu, Maokui He, Zide Liu, Xiyue Zhang, Xianwei Mao, Chunpeng Zhou, Jia Shi, Yanze Xin, Jingwen Li, Jingxie Zheng, Sijie Zeng, Chenfeng Wang, Fan Lu, Zeyu Zhang, Shuai Guo, Hengxuan Zhang, Pengfei Yu, Jia Shi, Yu Liu, Kun Zhan, Yan Xie

この論文をやさしく読む

ひとことで言うと

動作の経験を記憶にまとめ、次の判断や実行に利用して、モデルの再学習なしに改善を図るシステムです。

何に役立つ?

一度学習した能力を固定するだけでなく、蓄積した経験を再利用する身体性エージェントの設計に役立ちます。

この研究の面白いところ

経験の保存、スキルへの変換、行動生成を分け、判断に重要な時点や領域へ計算を集中させます。モデルの重みを再学習することとは異なる更新を扱います。

どこまで分かった?

評価は挙げられたベンチマークで、ME-RealBenchは6タスクです。スコアと成功率は別指標で、報告された差も相対的な増加率ではなくポイント差です。無制限の長期自己改善を実証したものではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

現在の身体性システムは主として、導入後も固定された事前学習済み能力に依存し、物理的な相互作用から学ぶ能力が制約されている。本研究では、行動実行、経験獲得、経験の進化、改善された実行という閉ループを中心に構成する、自己進化型の身体性システムMachEmbodied-Brain(ME-Brain)を導入する。Evolvable Memoryは、マルチモーダルな軌跡を階層的で再利用可能な経験へ統合する。Cognitive Coreは、物理的な経験を転用可能なスキルへ変える。Action Modelは、イベント駆動のキーフレーム、EventCellによる局所世界の予測、行動条件付きの記憶変調を組み合わせ、判断上重要な時点、領域、過去の証拠に計算を集中させる。これらのモジュールによって、モデルの再学習なしで、学習して固定する方式から、導入後に進化する方式へ身体性知能を移行させる。 Cognitive Coreは、身体性およびエージェントのベンチマークで、最も強い比較モデルをそれぞれ8.2ポイントと9.6ポイント上回った。Action ModelはRoboMMEで平均成功率47.88%を達成し、最も強いベースラインより3.26ポイント改善した。RoboDojoでは平均スコア21.51、成功率16.03%を達成し、π₀.₅をそれぞれ10.10ポイントと9.12ポイント上回った。6タスクのME-RealBenchでは、ME-Brainは平均スコア69.5、成功率66.7%を達成し、DM0.5をそれぞれ12.8ポイントと11.7ポイント上回った。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Current embodied systems largely rely on pretrained capabilities that remain fixed after deployment, limiting their ability to learn from physical interaction. We introduce MachEmbodied-Brain (ME-Brain), a self-evolving embodied system organized around a closed loop of action execution, experience acquisition, experience evolution, and improved execution. Evolvable Memory consolidates multimodal trajectories into hierarchical, reusable experience; Cognitive Core transforms physical experience into transferable skills; and the Action Model combines event-driven keyframes, EventCell local-world prediction, and action-conditioned memory modulation to focus computation on decision-critical moments, regions, and historical evidence. Together, these modules shift embodied intelligence from train-and-freeze to deploy-and-evolve without model retraining. Cognitive Core outperforms the strongest comparison models by 8.2 and 9.6 points on embodied and agent benchmarks. The Action Model achieves 47.88% mean success on RoboMME, a 3.26-point improvement over the strongest baseline. On RoboDojo, it reaches a 21.51 mean Score and 16.03% success rate, exceeding $\pi_{0.5}$ by 10.10 and 9.12 points. On the six-task ME-RealBench, ME-Brain achieves a 69.5 mean Score and 66.7% success rate, outperforming DM0.5 by 12.8 and 11.7 points, respectively.

arXiv ID: 2609.24271 / 要約の誤りについて