arXiv論文メモ
新着一覧
stat.ML / cs.LG · 査読状況未確認

解釈可能な機械学習で特徴量の関連性を分解する

Null importance: Disentangling relevance for interpretable machine learning

Garvesh Raskutti, Kris Sankaran, Jiaxin Ye

この論文をやさしく読む

ひとことで言うと

機械学習である特徴が重要でないとは何を意味するかを、統計的関連、予測性能、関数の変化、因果効果に分けて整理します。同じ重要度ゼロでも答えている問いが違う場合を扱います。

何に役立つ?

特徴重要度の結果から何を結論してよいか判断するために役立ちます。公平性やゲノム摂動のモデルでは、どの意味の関連を問うかによって解釈が変わることを説明します。

この研究の面白いところ

重要性の数値の大きさより先に、集団レベルでの無関連を定義します。異なる無関連が一致する十分条件と、一致しない反例を示し、各評価法が実際に何を測るかを結び付けます。

どこまで分かった?

理論上の同値性にはデータとモデルの仮定が必要です。シミュレーションや画像・マルチオミクスの例は区別の重要性を示しますが、一つの重要度手法ですべての科学的な問いに答えられるという結論ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

特徴量の重要度は解釈可能な機械学習の中心だが、「重要度」という語は、互いに大きく異なる関連性の概念を含んでいる。本研究では、特定した関連性の概念の下で、特徴量が無関係である条件を母集団レベルで特徴付けるnull importanceに基づき、統一的な見方を展開する。 限界的および条件付きの統計的関連性、予測リスク、関数的不変性、因果効果から生じる標準的なnull importanceを扱い、これらの概念が異なる科学的問いに答えることを示す。特に区別が重要になる二つの応用、すなわち、一般的な公平性基準が異なるnull importanceの概念に対応するアルゴリズム公平性と、関連性の概念が予測モデルの学習内容について異なる結論を導くゲノム摂動モデリングで枠組みを例示する。 この枠組みは、関連性を定義する科学的問い、異なるnullの関係を形作るデータとモデルの仮定、重要度を評価する方法という、特徴量解析の三側面を接続する。nullの概念が一致する十分条件を示し、それらの条件が成り立たないときにnullが分岐する反例を与える。さらに、どの手法群がどのnullを対象とするか、ゼロ重要度の統計量がその対象を特定するのはいつかを特徴付ける。 特徴量の依存、冗長性、非線形性、隠れた特徴など標準的な現象を含むシミュレーションと、画像・マルチオミクスデータの事例研究によって、これらの理論的な区別と実際の帰結に対する経験的証拠を示す。総合すると、本研究は科学的な問い、データ生成の仮定、アルゴリズムを関係付ける共通の統計言語を提供し、特徴量重要度の解析からどのような結論を支持できるかを明確にする。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-16(UTC)
最新改訂
2026-09-16 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Feature importance is central to interpretable machine learning, but the term "importance" encompasses several fundamentally different notions of relevance. We develop a unified perspective based on null importance: a population-level characterization of when a feature is irrelevant under a specified notion of relevance. We consider standard notions of null importance arising from marginal and conditional statistical relevance, predictive risk, functional invariance, and causal effects, and show how these notions answer different scientific questions. We illustrate the framework in two applications in which the distinction is particularly consequential: algorithmic fairness, where common fairness criteria correspond to different notions of null importance, and genomic perturbation modeling, where different notions of relevance lead to different conclusions about what a prediction model has learned. The framework connects three aspects of feature analysis: the scientific question defining relevance, the data and model assumptions that shape how different null notions relate, and the methods used to assess importance. We establish sufficient conditions under which null notions coincide and give counterexamples showing how they diverge when those conditions fail. We then characterize which nulls different method families target and when their zero-importance statistics identify those targets. Finally, simulations spanning feature dependence, redundancy, nonlinearity, hidden features and other standard phenomena, along with case studies on image and multiomics data, provide empirical evidence for these theoretical distinctions and their practical consequences. Taken together, these results provide a common statistical language for relating scientific questions, data-generating assumptions, and algorithms, and clarify the conclusions that feature-importance analyses can support.

著者のコメント

29 pages, 9 files. Submitted to Statistical Science

arXiv ID: 2609.19511 / 要約の誤りについて