世界モデルは過去の知識を忘れず組み合わせられるか
Benchmarking World Models for Continual Learning on Compositional Tasks
この論文をやさしく読む
ひとことで言うと
ロボットが覚えた動きや見え方を新しい組み合わせに使えるかを、単に新しい課題を学ぶ速さとは分けて評価します。
何に役立つ?
継続学習の改善が、既存知識の再利用によるものか、新規学習の能力によるものかを見分ける評価に役立ちます。
この研究の面白いところ
課題の組み合わせを行動と知覚に分けて設計しています。知識の再利用に向く部品を持つモジュール型モデルとも比較しています。
どこまで分かった?
モジュール型でも忘却と再利用の問題を完全には解決していません。要旨には個別の性能値や実機での検証条件は記載されていません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
世界モデルに望まれる性質の一つは、エージェントが既に学んだことを忘れずに、新しい環境に適応し、課題をまたいで継続的に学習できることである。とくに、物理世界の動力学は繰り返し現れる仕組みで記述できることが多いため、過去の経験から得た知識を保持し再利用する能力が、新しい環境へ効率よく適応する能力を支える。しかし、新たに来る課題には未知の内容と再登場する内容が混在するため、世界モデルの適応の測定では、未知課題を学ぶ速度と能力、既に得た知識の再利用という2つの能力が絡み合う。 過去の経験からの知識再利用を切り分けるため、ロボット操作の世界モデルを対象に、組合せ的な継続学習ベンチマークを提案する。具体的には、各課題系列に、それ以前に見た課題の側面を組み合わせた課題を設ける。さらに、その組み合わせを行動と知覚の軸に分け、異なる入力様式が知識再利用のどの部分を制約するかを理解しやすくする。代表的な継続学習手法の下で最先端の世界モデルを評価し、動力学の基盤に明示的に再利用可能な構成要素を含むモジュール型の世界モデルとも比較する。 結果は、モジュール化が従来手法より知識再利用と忘却のバランスをよく取ることを示すが、いずれの手法も問題を完全には解決せず、忘れずに再利用するよう作られた継続世界モデルには明確な改善余地が残る。詳細はプロジェクトのウェブサイトhttps://object814.github.io/Compositional-Continual-Learning/で公開されている。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-18(UTC)
- 最新改訂
- 2026-09-18 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-18 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
A desirable property of a world model is the ability to learn continually across tasks, adapting to new environments without forgetting what the agent has already learnt. In particular, the ability to retain and reuse knowledge obtained from prior experiences underpins an agent's ability to efficiently adapt to novel environments, as the dynamics of the physical world can often be described in recurring mechanisms. However, the world model's measure of adaptation entangles two abilities: the speed and capacity to learn unseen tasks, and the reuse of knowledge already acquired, since incoming tasks carry novel content alongside what recurs. In order to isolate knowledge reuse from prior experiences, we propose a compositional continual learning benchmark for world models in robot manipulation. Specifically, we design each task curriculum with compositional tasks that combine aspects of the tasks seen in the sequence. We further factorise this composition along the axes of action and perception to better understand how different input modalities bottleneck knowledge reuse. We evaluate state-of-the-art world models under canonical continual learning methods, alongside a modular world model whose dynamics backbone contains explicitly reusable components. Results show that modularity balances reuse against forgetting better than conventional methods, but none solve the problem fully, leaving clear room for continual world models built to reuse without forgetting. More details are available on our project website: https://object814.github.io/Compositional-Continual-Learning/.
arXiv ID: 2609.22055 / 要約の誤りについて