スパイクの競争で専門モデルを選ぶSpikeMoE
SpikeMoE: Brain-Inspired Competitive Routing for Flexible Spiking Mixture-of-Experts
この論文をやさしく読む
ひとことで言うと
発火回数の競争で使う専門モデルを選び、必要な部分だけ計算するSNNとMoEの統合方式です。
何に役立つ?
画像や言語を扱う計算で性能とエネルギー効率を両立し、入力の一部が欠ける場合にも対応する用途が考えられます。
この研究の面白いところ
専門モデルを選ぶ仕組み自体に側方抑制や不応期を組み込み、欠けた入力の表現を補う仕組みも追加しています。
どこまで分かった?
性能とエネルギー効率の良好な関係を述べていますが、要旨には消費電力の値や実機計測か推定かの説明はありません。比較対象に対する結果を、すべてのSNNやANNでの優位性とは解釈できません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
スパイキングニューラルネットワーク(SNN)は、神経細胞のスケールで生物に着想を得たダイナミクスを通じてイベント駆動計算を可能にする。一方、Mixture-of-Experts(MoE)は、モデルのスケールで専門モデルを選ぶことで条件付き計算を行う。両者の強みの統合は、柔軟なニューラル構成につながる可能性がある。ただし、スパイク活動に基づく専門モデル選択機構の設計が主要な課題である。 この課題に対し、海馬CA1領域で見られる競争・抑制に着想を得た、スパイクに基づくk-WTAルーターを導入する。このルーターは側方抑制と不応期を組み込み、離散的なスパイク数に従って上位K個の専門モデルを選ぶ。これを基盤として、神経細胞スケールのスパイクダイナミクスと、モデルスケールの専門モデル選択を統合する枠組みSpikeMoEを提示する。 マルチモーダル課題で感覚入力が不完全な場合に対処するため、二段階の欠落モダリティモデル化モジュールも備える。このモジュールは、観測済みモダリティのプールから得た経験的なプロトタイプと、モダリティ固有の学習可能な埋め込みを組み合わせ、欠落モダリティの表現を構成する。画像、言語、マルチモーダルのベンチマーク実験は、SpikeMoEがSNNの比較手法の中で最先端の性能を達成し、通常の人工ニューラルネットワーク(ANN)の対応モデルと同等以上の性能を持ち、多様なモダリティ欠落条件でも頑健性を保つことを示す。これらの結果は、性能とエネルギー効率の間の良好なトレードオフを示し、スパイクダイナミクスと疎な専門モデル計算の統合を裏付け、エネルギー効率の高い脳型計算の有望な手法としてSpikeMoEを示している。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Spiking Neural Networks (SNNs) enable event-driven computation through biologically inspired dynamics at the neuronal scale, while Mixture-of-Experts (MoE) perform conditional computation through expert selection at the model scale. Integrating their strengths offers potential for flexible neural architectures. A key challenge, however, lies in designing an expert selection mechanism based on spiking activity. To address this, we introduce a spike-based k-WTA Router inspired by competition-inhibition observed in the hippocampal CA1 region. The router incorporates lateral inhibition and refractory period to select Top-K experts according to discrete spike counts. Building on this, we present SpikeMoE, a framework that integrates neuronal-scale spiking dynamics with model-scale expert selection. To address incomplete multisensory inputs in multimodal tasks, we further equip SpikeMoE with a two-stage missing-modality modeling module that combines empirical prototypes from an observed-modality pool with modality-specific learnable embeddings to construct missing-modality representations. Experiments on vision, language, and multimodal benchmarks demonstrate that SpikeMoE achieves state-of-the-art performance among the SNN baselines, matches or exceeds the performance of ANN counterparts, and maintains robustness across diverse missing-modality conditions. These results demonstrate a favorable trade-off between performance and energy efficiency, validating the integration of spiking dynamics with sparse expert computation and highlighting SpikeMoE as a promising approach to energy-efficient brain-inspired computing.
arXiv ID: 2610.01418 / 要約の誤りについて