arXiv論文メモ
新着一覧
cs.LG / q-bio.NC · 査読状況未確認

139匹のマウスの脳記録で汎用表現の転用能力を評価

BrainWideBench: Benchmarking large-scale pretraining and across-animal transfer in multi-region neural recordings

Alexandre Andre, Shivashriganesh P. Mahato, Vinam Arora, Keshav Balaji, Divyansha Lachi, Nanda H. Krishna, Jingyun Xiao, Yizi Zhang, Ximeng Mao, Wenrui Ma, Han Yu, International Brain Laboratory, Daniel Birman, Niccolò Bonacchi, Gaelle A. Chapuis, Joana A. Catarino, Felicia Davatolhagh, Mayo Faulkner, Laura Freitas-Silva, Fei Hu, Julia M. Huntenburg, Anup Khanal, Inês Laranjeira, Petrina Lau, Guido T. Meijer, Nathaniel J. Miska, Jean-Paul Noel, Alejandro Pan-Vazquez, Georg Raiser, Cyrille Rossant, Karolina Z. Socha, Anne E. Urai, Miles J. Wells, Steven J. West, Olivier Winter, Blake Richards, Guillaume Lajoie, Cole Hurwitz, Mehdi Azabou, Matthew R. Whiteway, Liam Paninski, Eva L. Dyer

この論文をやさしく読む

ひとことで言うと

別々のマウスから学んだ脳活動のモデルが、新しい個体でも行動や神経活動、脳領域の構成を捉えられるかを共通の基準で比較します。

何に役立つ?

脳活動モデルの事前学習が何に役立ち、どの課題には転用しにくいかを調べる評価に役立ちます。

この研究の面白いところ

行動を当てる能力だけでなく、神経活動の予測と解剖学的な構成の復元も同時に評価します。事前学習が得意な課題に偏っていないかが見えます。

どこまで分かった?

対象は特定の意思決定課題を行うマウスの記録です。人間の脳や任意の行動への一般化を示したものではなく、3課題群すべてで良い手法も未達です。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

大規模な神経記録技術の進歩により、多数の動物や離れた脳領域からデータを収集できるようになり、この規模を活用して、多様な下流課題へ転用できる汎用的な神経表現を学べるかが問われている。しかし、その進展は、評価手順の分断と、個別の課題領域への狭い焦点によって制限されてきた。そこで、複数脳領域の神経記録における動物個体間の転用を評価するベンチマークBrainWideBenchを提案する。International Brain LaboratoryのBrainwide Mapデータセットを基にし、感覚に導かれた意思決定課題を行う139匹のマウスから記録した、276脳領域にわたる神経・行動データを用いる。 ベンチマークは、学習した表現が下流の行動復号を支えられるか、隠した神経活動や将来の神経活動を予測できるか、生物学的に意味のある解剖学的構成を復元できるかを調べる、相補的な3つの課題群からなる。このベンチマークで、下流目的に合わせた微調整や未見の動物へのゼロショット一般化を含む転用設定にわたり、事前学習手法を系統的に評価する。 結果は、条件をそろえた単一セッションの基準モデルより事前学習が性能を改善することを確認する一方、現在の手法の転用能力にはばらつきがあり、改善は事前学習の目的と下流課題の一致に強く依存することを示す。3課題群すべてで一様に良い単一の手法はなく、大半の手法はその一部だけを扱うよう設計されている。これらの知見は、行動、動力学、解剖学のすべてへ一般化する表現の学習が未解決であることを示唆する。BrainWideBenchは統一された再現可能な評価群を提供し、マウスの脳の汎用モデルへ向けた進展を測る枠組みを確立する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Advances in large-scale neural recording have made it possible to collect data across many animals and distributed brain regions, raising the question of whether this scale can be exploited to learn general-purpose neural representations transferable across diverse downstream tasks. Yet, progress toward this goal has been limited by fragmented evaluation protocols and a narrow focus on individual task domains. Here, we present BrainWideBench, a benchmark for evaluating across-animal transfer on multi-region neural recordings, built on the International Brain Laboratory Brainwide Map dataset of neural and behavioral recordings spanning 276 brain regions from 139 mice performing a sensory-guided decision-making task. The benchmark is organized around three complementary task suites that evaluate whether learned representations support downstream decoding of behavior, can predict masked or future neural activity, and can recover biologically meaningful anatomical organization. With this benchmark, we systematically evaluate pretraining methods across transfer settings, including finetuning on downstream objectives and zero-shot generalization to unseen animals. Our results confirm pretraining improves performance over matched single-session baselines, but we show current methods exhibit heterogeneity in transfer capabilities: gains depend strongly on the alignment between pretraining objectives and downstream tasks. No single approach performs uniformly well across all three suites, and most methods are designed to only address a subset of them. Together, these findings suggest that learning representations that jointly generalize across behavior, dynamics, and anatomy remains an open challenge. By providing a unified and reproducible evaluation suite, BrainWideBench establishes a framework for measuring progress toward general-purpose models of the mouse brain.

arXiv ID: 2609.22064 / 要約の誤りについて