arXiv論文メモ
新着一覧
cs.CV · 査読状況未確認

42拠点の脳MRIを連合事前学習する基盤モデルBrainFedFM

A generalizable structural brain MRI foundation model built through dual-priority federated pretraining

Zhen Yu, Yang Liu, Xiahai Zhuang, Qingchao Chen

この論文をやさしく読む

ひとことで言うと

42拠点の脳MRI画像を集約せずに学習し、多様な解析課題へ使える基盤モデルを作った。

何に役立つ?

プライバシーと管理の制約がある複数拠点のMRIデータを使うモデル開発に役立つ可能性がある。

この研究の面白いところ

各拠点の重要な脳領域と、小規模拠点を含む全体への寄与を別々に優先付けした。

どこまで分かった?

20データセット、17課題、七モデルとの比較結果である。臨床判断での有効性や各拠点での実運用を直接保証するものではない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

基盤モデルは、発達、加齢、疾患にわたる構造的脳MRIの汎用的な解析に有望である。しかし従来のモデルは、プライバシーやデータ管理の制約があるにもかかわらず、データを一か所に集めて事前学習することが多い。そのような最適化は集団の規模を重視しすぎ、小規模で専門的な集団の補完的情報を見落とし得る。本研究は、多様な実世界のデータ分布から得た三次元画像164,707件を42の連合拠点に分けて事前学習した、構造的脳MRIの基盤モデルBrainFedFMを提示する。 BrainFedFMは二重の優先付けを伴う連合事前学習を用いる。各拠点での空間的な優先マスキングと、サーバーでの拠点優先の集約を結び付け、局所では情報量の高い解剖学的領域を、全体では拠点ごとの寄与を重視する。分類、回帰、領域分割の17種類の課題を含む20の下流データセットで、中央集約型基盤モデル四つを含む七モデルの中で最高水準の性能を達成し、平均順位は1.68、改善率は50%だった。特に分類と回帰で一貫した優位性があり、十分に代表されない集団に対しても頑健だった。この結果はBrainFedFMの汎用性を示し、生の画像を集めずに分散データから神経画像の基盤モデルを開発する実用的な方法として、連合事前学習を位置付ける。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Foundation models hold promise for generalizable analysis of structural brain magnetic resonance imaging (MRI) across development, aging and disease. However, existing models are typically built through centralized pretraining on pooled data, despite privacy and governance constraints. Such pooling optimization can overemphasize cohort size and overlook complementary information from smaller, specialized cohorts. Here we present BrainFedFM, a structural brain MRI foundation model federatively pretrained on 164,707 three-dimensional scans drawn from diverse real-world data distributions and organized across 42 federated sites. BrainFedFM uses dual-priority federated pretraining, coupling spatial-priority masking at each site with site-priority aggregation at the server to emphasize informative anatomical regions locally and prioritize site contributions globally. Across 20 downstream datasets spanning 17 classification, regression and segmentation tasks, BrainFedFM achieved the state-of-the-art performance (mean rank 1.68, 50\% gain) across seven models, including four centralized foundation models, while showing particularly consistent advantages in classification and regression and robustness across underrepresented populations. These findings demonstrate the generalizability of BrainFedFM and highlight federated pretraining as a practical strategy for developing neuroimaging foundation models from distributed data without pooling raw images.

arXiv ID: 2609.27611 / 要約の誤りについて