arXiv論文メモ
新着一覧
stat.ML / cs.AI / cs.LG / math.ST / stat.ME / stat.TH · 査読状況未確認

予測値の区間ごとに較正する予測区間

PICPIs: Prediction-Interval-Conditional Prediction Intervals

Xuelin Yang, Baihe Huang, Yilong Hou, Guido Imbens, Michael I. Jordan

この論文をやさしく読む

ひとことで言うと

予測値を区間に分け、その区間に入る事例の結果の平均も同じ区間に入るようにする不確実性評価法です。

何に役立つ?

考えられる用途は、予測値の範囲ごとの較正状況を見ながら意思決定することです。個々の事例の結果を必ず区間に含めるという保証ではありません。

この研究の面白いところ

元の予測値を変更せず、区間自身を条件づけの単位と保証の対象にする自己整合条件を導入しています。

どこまで分かった?

区間幅の収束結果には予測分布の正則性などの条件が必要で、予測誤差と対数因子も残ります。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

統計学の古典的な問題の一つは、観測できない対象を推論するとき、観測可能な量の何を条件にすべきかである。非パラメトリックな不確実性評価に用いる適合予測では、標準的な周辺的な妥当性だけでは、意思決定の基になる予測値ごとの情報が粗い。一方、共変量に関する完全な条件付き保証は、原理的に達成できないことが知られている。 本研究はこの間を埋めるため、Prediction-Interval-Conditional Prediction Intervals(PICPI)と呼ぶ、予測値に基づく条件づけの枠組みを導入する。形式的には、予測モデル p、状況を表す共変量 X、結果 Y に対し、区間 I が E[Y|p(X)∈I]∈I という自己整合条件を満たすときPICPIとする。つまり区間は、予測値の一群を定めると同時に、その群の結果の平均が同じ区間内にあることを保証する。この条件は元の予測を変えずに、データに適応した区分を与える。実用的なアルゴリズムでそのような区間を構成できる。 予測値の分布に正則性があれば、構成した区間は予測値の任意に小さくできる割合を除いて全体を覆い、幅は対数因子と予測誤差を除いて n^(-1/3) の速さで縮む。局所的に較正されたこれらの区間を特定することは、後段の意思決定にも役立ち得る。確率的予測と多クラス分類について、理論的保証を伴うPICPIの推論手続きを導き、既存の区間型手法と比較する実証結果も示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

A classical question in statistics is which observable quantities to condition on when drawing inferences about unobservable targets. For conformal prediction in nonparametric uncertainty quantification, standard marginal validity offers limited resolution at the prediction values on which decisions are based, and fully conditional guarantees with respect to the covariates are provably unattainable. We address this gap by introducing a prediction-based conditioning framework that we refer to as Prediction-Interval-Conditional Prediction Intervals (PICPIs). Formally, a PICPI is an interval $I$ satisfying a self-consistency condition: $$\mathbb{E} [Y \mid p(X) \in I] \in I,$$ for predictive model $p$, contextual covariate $X$, and outcome $Y$. Thus, an interval simultaneously defines a stratum of prediction values and certifies that the mean outcome in that stratum lies in the same interval. This self-consistency condition yields data-adaptive strata without altering the original prediction. Such intervals can be constructed using practical algorithms. Under regularity of the prediction distribution, the constructed intervals cover all but an arbitrarily small fraction of prediction values and have widths that decrease at rate $n^{-1/3}$, up to logarithmic factors and the prediction error. Moreover, identifying these locally calibrated intervals can, in turn, inform downstream decision-making. We derive inference procedures for PICPIs in probabilistic prediction and multi-class classification, accompanied by theoretical guarantees. Empirical results are provided that compare PICPIs with existing interval-based baselines.

著者のコメント

45 pages, 6 figures

arXiv ID: 2609.25388 / 要約の誤りについて