まれな環境変化に備えるCVaRアンカー回帰
CVaR anchor regression protects against rare shifts
この論文をやさしく読む
ひとことで言うと
普段の予測精度を保ちながら、まれに起きる大きな環境変化にも備える回帰方法。
何に役立つ?
学習時に少数の大きな環境変化がある予測問題で、頑健さと通常時の精度の調整に役立つ。
この研究の面白いところ
環境ごとの残差の端の部分だけを平均し、まれな変化への重み付けを調節できる。
どこまで分かった?
厳密なリスク保証は、分散が変わり得る線形構造モデルの仮定の下でのもの。実データへの適用例はニューヨーク市のタクシー。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
学習データにまれだが大きな変化が含まれるとき、新しい環境での予測を研究する。アンカー回帰は、環境ごとの平均残差の二乗を平均した値に罰則を与え、学習時の変化の二次モーメントで決まる楕円体の範囲内の変化に備える。そのため、まれな変化を覆うには強い罰則が必要になり、楕円体があらゆる方向に広がって、通常の環境での精度を下げる可能性がある。そこで、平均残差の二乗の単純平均を、分布の端の平均へ置き換えるCVaRアンカー回帰を提案する。予測のリスクへCVaRやGroupDROを直接適用する場合とは異なり、雑音が大きいというだけで環境の重みを増やさない。 雑音の分散が環境によって異なり得る線形構造モデルの下で、最悪の場合のリスクについて正確な保証を証明する。環境が離散的な場合、CVaRで平均する分布の端の割合を小さくすると、頑健性の対象集合は楕円体から、学習時の変化とその符号を反転したものの凸包を拡大した集合へ広がる。その拡大率は別のパラメータで制御する。例により、通常の環境での精度を保ちながら、まれな変化への備えを改善できることを示す。また、ニューヨーク市のタクシーデータで方法を例示する。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-22(UTC)
- 最新改訂
- 2026-09-22 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-22 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We study prediction in new environments when training data contain rare, large shifts. Anchor regression penalizes the average of the squared mean residual across environments. It protects against shifts in an ellipsoid determined by the second moment of the training shifts. Covering rare shifts may therefore require a large penalty, expanding the ellipsoid in every direction and reducing accuracy on common environments. We propose CVaR anchor regression, which replaces the average of the squared mean residuals with a tail average. Unlike CVaR or GroupDRO applied directly to prediction risks, it does not give environments more weight solely because their noise levels are high. We prove an exact worst-case risk guarantee under a linear structural model that allows for heteroscedastic noise. For discrete environments, decreasing the CVaR tail fraction expands the robustness set from an ellipsoid to a scaled convex hull of the training shifts and their negatives. A separate parameter controls its scale. Examples show how the method can improve protection against rare shifts while retaining accuracy on common environments. We illustrate the method on New York City taxi data.
arXiv ID: 2609.27034 / 要約の誤りについて