arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

複数の予測対象を結ぶ効果を残差の相関から見積もる

Residual Correlation as a Diagnostic for Joint-Uncertainty Gains from GP Coregionalisation

Fangqin Zhou and Joaquin Vanschoren

この論文をやさしく読む

ひとことで言うと

複数の予測対象を結ぶと不確実性の推定が改善するか、独立モデルの残差から見積もった。

何に役立つ?

複数対象の予測モデルを結合するかどうかを判断する診断に役立つ。点予測の改善を示す結果ではない。

この研究の面白いところ

対象同士の元の相関より、個別の予測後に残る相関のほうが同時不確実性の改善をよく示した。

どこまで分かった?

16件のベンチマークなどで診断量とNLL改善の強い関連を報告した。診断量は大域的なガウス型残差依存と分離可能な共地域化に特化している。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

複数の対象を予測する回帰では、相関する対象を固有の共地域化モデルを持つ多出力ガウス過程(GP-ICM)で結び、情報を共有すれば全体の性能が上がると想定される。しかし実際の利点は一定しない。本研究の設定では、共地域化の主な利点は点予測ではなく、複数対象を合わせた不確実性の定量化にあった。対象間の生の相関は結合が役立つ時期を予測しない。ここで調べた分離可能なGP-ICMの設定では、対象ごとに独立した予測器で説明できずに残る対象間の依存、すなわち残差の相関が、同時不確実性の改善を最もよく予測した。独立したGPだけから計算できる軽量な診断量D_logdet=−(1/2)log det R_resを導入する。これは、対角の残差共分散の代わりに完全な共分散をモデル化した場合の理想化された同時負対数尤度(NLL)の改善を表す。制御された合成データの研究、16件の複数対象ベンチマーク、キーポイント回帰向けの固定されたTransformerと畳み込みニューラルネットワークの表現で、点予測はほぼ変わらなかった(ΔR²≈0)。一方、この診断量は観測されたICMのNLL改善を強く予測し(スピアマンのρ=−0.83、p<0.001)、特徴量数とサンプル数の比などの経験則より優れた。また、各対象の周辺分散を独立モデルのまま保ち、同時共分散に残差相関の構造を加えるResidual-ICMを提案する。比較した方法の中でResidual-ICMは平均の同時NLLが最良となり、診断量は共分散を結合すると有用そうな場合を示した。この診断量が対象とするのは、分離可能な共地域化が捉える大域的なガウス型の残差依存である。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

In multi-target regression, correlated targets are often coupled through multi-output Gaussian processes with an intrinsic model of coregionalisation (GP-ICM), assuming that sharing statistical strength improves overall performance. In practice, the benefits are inconsistent. Across the settings studied, we find that the main benefit of coregionalisation is joint uncertainty quantification rather than point prediction. Raw target correlation does not predict when coupling helps; in the separable GP-ICM settings studied here, residual correlation, the cross-target dependence left unexplained by independent per-target predictors, is the strongest predictor of joint-uncertainty gains. We introduce a lightweight diagnostic, $D_{\rm logdet}=-\frac{1}{2}\log\det R_{\rm res}$, which represents the idealised joint negative log-likelihood (NLL) gain from modelling a full rather than diagonal residual covariance and is computable from independent GPs alone. Across a controlled synthetic study, 16 multi-target benchmarks, and frozen transformer and convolutional neural network representations for keypoint regression, point prediction remains largely unchanged ($\Delta R^2\approx 0$). In contrast, $D_{\rm logdet}$ strongly predicts observed ICM NLL improvements ($\rho_s=-0.83$, $p<0.001$), outperforming heuristics such as the feature-to-sample ratio. We also propose Residual-ICM, which preserves independent marginal variances while adding residual-correlation structure to the joint covariance. Residual-ICM achieves the best average joint NLL among the compared methods, while the diagnostic indicates when covariance coupling is likely to be useful. The diagnostic is specific to global Gaussian residual dependence, the structure captured by separable coregionalisation.

著者のコメント

Accepted at ACML 2026

arXiv ID: 2609.30085 / 要約の誤りについて