arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

潜在表現の整合性から予測の過信を抑える

Robust Evidential Learning Through Latent Consistency

Charmaine Barker, Daniel Bethell, Simos Gerasimou

この論文をやさしく読む

ひとことで言うと

予測の答えはそのままに、モデル内部の表現が不安定な入力について自信だけを下げる手法です。

何に役立つ?

学習済みモデルを再学習せず、不確実性の評価を改善する用途が考えられます。実証は要旨に挙げられた画像関連などのベンチマークによるものです。

この研究の面白いところ

入力画像を何度も処理する代わりに、内部の潜在空間で摂動を与えて整合性を確認するため、事後処理の速度と予測性能の維持を両立させています。

どこまで分かった?

要旨のAUROC改善値+8.29、+5.01には尺度の詳細が明記されていません。較正データを必要とし、すべての未知入力や攻撃への保証が示されているわけではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

学習分布外の入力や敵対的入力によって、自信は強いのに信頼できない予測が生じ得るため、重大な影響を伴う場面で深層学習モデルを導入するには、信頼できる不確実性の定量化が不可欠である。証拠に基づく深層学習は1回の順伝播で効率的に不確実性を推定できるが、敵対的入力など、学習済み表現による裏付けが乏しい入力にも高い証拠強度を与えてしまうことがある。 そこで、再学習も元の予測の変更も行わずに証拠の頑健性を改善する、軽量でタスクに依存しない事後処理手法CLEARを提案する。CLEARは、学習に使わずに確保した較正データを用いて、モデルの潜在空間におけるグループごとの幾何学的構造を特徴付ける。推論時には、潜在空間内で直接、摂動を加えた複数の表現を効率よく生成し、予測されたグループの較正済み幾何構造に照らして、それらの不整合を測定する。潜在空間での不整合が大きいことは、証拠に裏付けがないことを示す。CLEARはこれを用いて証拠強度を選択的に下げる一方、潜在表現が整合的な入力については証拠を保持する。ImageNet→CUBの設定で、CLEARは分布外入力と敵対的入力に対するAUROCをそれぞれ+8.29、+5.01改善する。また、競合する事後処理手法より17.4倍高速に動作し、分類、回帰、物体検出の各ベンチマークで予測性能を維持する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-10-01(UTC)
最新改訂
2026-10-01 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Reliable uncertainty quantification is essential for deploying deep learning models in high-stakes settings, where out-of-distribution and adversarial inputs can induce confident but unreliable predictions. Evidential Deep Learning provides efficient uncertainty estimates in a single forward pass, but can still assign high evidential strength to inputs that are poorly supported by the learned representation, such as adversarial inputs. We introduce CLEAR, a lightweight, task-agnostic post-hoc method that improves evidential robustness without retraining or altering the base prediction. Using held-out calibration data, CLEAR characterises the group-conditioned geometry of the model's latent space. At inference, it efficiently generates perturbation views directly in the latent space and measures their conflict relative to the calibrated geometry of the predicted group. High latent conflict indicates unsupported evidence, which CLEAR uses to selectively reduce evidential strength while retaining evidence for latent-consistent inputs. On ImageNet$\rightarrow$CUB, CLEAR improves OOD and adversarial AUROC by $+8.29$ and $+5.01$ while running 17.4$\times$ faster than competing post-hoc methods while preserving predictive performance across classification, regression, and object detection benchmarks.

arXiv ID: 2610.01384 / 要約の誤りについて