arXiv論文メモ
新着一覧
cs.LG / stat.ML · 査読状況未確認

特徴の符号と積を保つ幾何平均プーリング

Geometric Mean Pooling for Equal-Weight Multiplicative Coarse-Graining

Ang-Kun Wu, Fangdi Wen, Jingtao Zhang

この論文をやさしく読む

ひとことで言うと

複数の特徴をまとめる際に、足し算や最大値の代わりに、符号と大きさの積を生かす方法です。積の関係が重要な信号を失いにくくすることを狙っています。

何に役立つ?

考えられる用途は、等しい重みで特徴を掛け合わせる構造が自然な学習課題です。合成タスクでは利点が示されましたが、画像や分子では使い方によって効果が変わります。

この研究の面白いところ

量子多体系の局所から大域への合成をヒントに、追加の学習パラメータなしで符号情報も残します。重ならない階層構造で大域的な乗法統計量を保つ性質を示しています。

どこまで分かった?

著者自身が汎用的な置き換えではないと位置づけています。ノイズ耐性は試した水準での結果で、画像や分子での有効性は表現とプーリングの配置などに依存します。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

平均プーリングの加算的な偏りと最大値プーリングの極値への偏りに代わるものとして、特徴の符号の積と特徴の大きさの幾何平均を組み合わせる、符号付きプーリング演算子「幾何平均プーリング」(GMP)を導入する。量子多体系の物理における局所から大域への合成に着想を得たGMPは、学習可能なプーリングパラメータを導入せずに、符号の組合せに関する情報と特徴的な乗法的スケールの両方を保持する。 重なりのない階層的GMPが、対応する大域的な乗法統計量を保存することを示し、合成系列タスク、反復的な粗視化、画像分類、分子の脂溶性回帰で評価する。合成タスクでは、GMPは平均および最大値プーリングより正確に積に基づく信号を復元し、試した水準の乗法的な入力ノイズのもとで予測性能を維持する。一方、画像と分子のデータでは、その有効性は表現、予測対象のパラメータ化、局所・大域プーリングの配置に依存する。これらの結果は、GMPを標準的なプーリング演算子の普遍的な代替とするのではなく、等しい重みでの乗法的合成が妥当と考えられるタスクに対する、条件に依存した相補的な帰納バイアスとして位置づける。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

As an alternative to the additive and extremal biases of average and max pooling, we introduce Geometric Mean Pooling (GMP), a signed pooling operator that combines the product of feature signs with the geometric mean of feature magnitudes. Motivated by local-to-global composition in quantum many-body physics, GMP retains both joint sign information and a characteristic multiplicative scale without introducing learnable pooling parameters. We show that non-overlapping hierarchical GMP preserves the corresponding global multiplicative statistic and evaluate it on synthetic sequence tasks, iterative coarse-graining, image classification, and molecular lipophilicity regression. On the synthetic tasks, GMP recovers product-based signals more accurately than average and max pooling and maintains predictive performance under the tested levels of multiplicative input noise. On image and molecular data, however, its effectiveness depends on the representation, target parameterization, and placement of local and global pooling. These results position GMP as a complementary, regime-dependent inductive bias for tasks in which equal-weight multiplicative composition is plausible, rather than as a universal replacement for standard pooling operators.

著者のコメント

17 pages, 6 figures

arXiv ID: 2609.21876 / 要約の誤りについて