arXiv論文メモ
新着一覧
cs.LG / cs.AI · 査読状況未確認

脳の階層的な役割分担を参考にした継続学習

Brain-Inspired Hierarchical Modularity for General Continual Learning

Hongwei Yan, Kanglei Zhou, Qi Cheng, Weiyi Dong, Chunyan Lan, Guanglong Sun, Jun Zhou, Qian Li, Yi Zhong, Liyuan Wang

この論文をやさしく読む

ひとことで言うと

経験が次々に変わる中で、知識の干渉を抑えつつ新しいことを学ぶためのモジュール構成。

何に役立つ?

考えられる用途は、オンラインで変化するデータを扱う認識・ロボット学習。要旨では複数課題での評価を報告する。

この研究の面白いところ

衝突する経験を専門家に分け、両立する経験は空間・時間の複数尺度で統合する設計。

どこまで分かった?

50パーセントポイント超の改善は身体的な物体操作で、再生データを使わない代替手法との比較。すべての課題で同じ改善幅とは述べられていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

継続学習は、以前の知識を保持し適応させながら、順番に得られる経験から学ぶ能力であり、変化する環境で動く知能システムに重要である。しかし従来の研究は、課題の区切りが明確なオフラインの課題別学習を扱うことが多く、オンラインで不確実かつ変化するデータの流れにおける一般的な継続学習との隔たりが大きい。この状況では、干渉を減らすため衝突する経験を分離しながら、汎化を促すため両立する経験を統合する必要がある。 著者らはショウジョウバエの学習・記憶システムの構成から着想を得て、専門家の役割分担と集合的な統合を協調させる階層的なモジュール原理を見いだす。これを事前学習済み基盤モデルの軽量なモジュール適応として具体化し、専門家への振り分けに脳を参考にしたランダムな拡張を使い、空間・時間の複数の尺度で多様なモジュールを統合する。 画像認識、視覚・言語の理解、一人称・三人称動画の理解、身体的な視覚・言語・行動学習の各課題で、オンラインかつ不確実なデータの流れでの学習を一貫して改善した。身体的な物体操作では、再生データを使わない代替手法と比べて50パーセントポイントを超える改善を得た。著者らは、動的な経験からの学習に対して階層的モジュール性が生物学的根拠を持つ方向性を示すと述べる。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Continual learning, the ability to learn from sequential experience while retaining and adapting prior knowledge, is central to intelligent systems operating in changing environments. However, conventional continual learning is typically studied with offline task-wise training and clear task boundaries, leaving a substantial gap from general continual learning under online, uncertain, and evolving data streams. In this regime, intelligent systems must separate conflicting experience to reduce interference while integrating compatible experience to promote generalization. Inspired by the organization of the Drosophila learning and memory system, we identify a hierarchical modular principle that coordinates both functions through expert specialization and ensemble integration. We instantiate this principle as lightweight modular adaptation of pretrained foundation models, combining brain-inspired random expansion for expert routing and diversified modular integration across spatial and temporal scales. Across visual recognition, vision-language understanding, ego-exo video understanding, and embodied vision-language-action learning, our method consistently improves learning under online and uncertain data streams, with gains exceeding 50 percentage points over replay-free alternatives in embodied manipulation. These findings support hierarchical modularity as a biologically grounded path for learning from dynamic experience.

著者のコメント

50 pages

arXiv ID: 2609.25146 / 要約の誤りについて