arXiv論文メモ
新着一覧
cs.LG / cs.AI · 査読状況未確認

少数の特徴で予測を説明する確率的線形モデル

Probabilistic Linear Explanations

Frederic Koriche and Jean-Marie Lagniez and Chi Tran

この論文をやさしく読む

ひとことで言うと

予測理由を少数の特徴だけで表し、それぞれが予測をどちら向きにどれだけ動かすかを示す手法です。

何に役立つ?

分類と回帰の両方で、特徴数を制限した説明を作る用途があります。説明の簡潔さと元モデルとの対応を数学的に扱います。

この研究の面白いところ

本来の最適化が難しいことを示したうえで、代理誤差との関係を証明し、混合整数計画と反復法の2つの解法を用意しています。

どこまで分かった?

理論的な誤差関係は記載された局所分布族に対するものです。MIPの標本計算量が多項式であることと、求解時間が多項式であることは別です。要旨の計算量クラス表記は未展開の記号を含みます。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

形式的な説明可能性は、個々の予測に数学的根拠を持つ理由を与える。しかし、アブダクティブな説明は特徴数が多すぎて人間の認知的限界を超えることが多く、確率的な緩和も主としてカテゴリ分類に限られてきた。本研究では、疎で基準点に固定された線形モデルに基づき、二値分類と連続値回帰の両方に適用できる、確率的な説明可能性の統一的枠組みを示す。事例をブール超立方体へ写像することで、提案する線形説明は、部分集合に基づく手法を真に一般化する。指定された疎性の上限kを守りながら、特徴の寄与の大きさと方向の両方を捉える。 対象モデルがニューラルネットワークの場合、このような説明の関連性誤差を最小化することがNPPP困難であると示し、この扱いにくい目的を、扱いやすい代理目的である忠実度誤差と関連づける。パラメータ付きの局所分布族について、任意のk疎な説明の関連性誤差は、その忠実度誤差に、局所的には小さく保たれる乗法係数を掛けた値で抑えられる。得られる経験的な問題には、相補的な2つの方法で取り組む。1つは、多項式の標本計算量を維持しながら、経験的な問題に対して最適性を証明できる解を得る混合整数計画(MIP)の定式化である。もう1つは、証明可能な近似保証を持つ、多項式時間の反復ハードしきい値処理(IHT)アルゴリズムである。実証評価では、LIMEやMAPLEなどの先端的な比較手法と異なり、提案する説明は構成上、基準点への固定と疎性の制約の両方を満たし、同時に一貫して低い関連性誤差を達成する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-16(UTC)
最新改訂
2026-09-16 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Formal explainability provides mathematically grounded justifications for individual predictions. However, abductive explanations often exceed human cognitive limits by involving too many features, while probabilistic relaxations have remained largely limited to categorical classification. We present a unified framework for probabilistic explainability based on sparse, anchored linear models, applicable to both binary classification and continuous regression. By mapping instances to the Boolean hypercube, our linear explanations strictly generalize subset-based approaches: they capture both the magnitude and direction of feature contributions while enforcing a prescribed sparsity budget $k$. We show that minimizing the relevance error for such explanations is \ClassNPPP-hard when the underlying model is a neural network, and we relate this intractable objective to a tractable surrogate---the fidelity error. For a parameterized family of local distributions, the relevance error of any $k$-sparse explanation is bounded by its fidelity error up to a multiplicative factor that remains small locally. We address the resulting empirical problem using two complementary approaches: a Mixed Integer Programming (MIP) formulation that yields provably optimal empirical solutions while maintaining polynomial sample complexity, and a polynomial-time Iterative Hard Thresholding (IHT) algorithm with provable approximation guarantees. Empirical evaluations show that, unlike state-of-the-art baselines such as LIME and MAPLE, our explanations satisfy both the anchoring and sparsity constraints by construction, while consistently achieving lower relevance error.

著者のコメント

Under Review

arXiv ID: 2609.19077 / 要約の誤りについて