arXiv論文メモ
新着一覧
cs.LG / cs.CY · 査読状況未確認

機械学習の公平性制約が改善する場合と悪化する場合

When Post-Processing Fairness Constraints Help and When They Harm: Evidence from Eight Cross-Domain Evaluations

Nithin Raghava Ramachandra Narla

この論文をやさしく読む

ひとことで言うと

公平性を調整する後処理が、元の格差や測定方法によって改善にも悪化にもなることを調べた。

何に役立つ?

機械学習モデルを複数分野に導入するとき、公平性調整の適用判断と継続監視に役立つ。

この研究の面白いところ

格差が大きい14例のうち9例で改善した一方、ほぼ公平な4例のうち3例では悪化した。

どこまで分かった?

一つの後処理手法を8分野で評価した結果であり、すべての公平性手法に同じ傾向があるとは言えない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

実運用の機械学習での公平性監査は通常、導入時に一つの分野で一度だけ行われる。しかし再学習や利用者層の変化で公平性は変わり得るし、一つのデータで確認した介入は、組織が使う多様な分野で試されることが少ない。本研究は四段階の枠組みFAPEを提示し、FairlearnのThresholdOptimizerという一つの後処理介入を、刑事司法、所得予測、法学課程への入学、信用融資、農業融資、複数分野のベンチマーク、医療、教育の8分野で評価する。人口統計的な均等と均等化オッズの差に加え、計算できる場合は不均衡影響比と精度への負担も評価した。 介入の効果は元の格差の大きさと関係した。モデルと分野の組み合わせのうち、元の格差が大きい14例では9例で改善したが、元からほぼ公平な4例では3例で悪化した。格差が大きい例に残る5件の例外は、最小グループ人数か、保留データでしきい値を適合するかのいずれかの測定確認によって判定が逆転した。導入時に開始するCUSUM監視器を模擬的な変化で試すと、0.1という均等性の慣例的基準を一度も満たさない制約付きモデルと、当初は満たしたが後に悪化したモデルを区別できた。導入時だけの監査は信頼できる指針ではなく、元の格差による事前選別と継続的な監視が必要だと論じる。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Fairness audits in production ML typically occur once, at deployment, on a single domain. Both fail in practice: fairness can shift after retraining or a changing user base, and interventions validated on one dataset are rarely tested across the heterogeneous domains an organization deploys. We present FAPE (Fairness Auditing for Production Environments), a four-stage framework evaluating a single post-processing intervention, Fairlearn's ThresholdOptimizer, across eight domain evaluations: criminal justice, income prediction, legal admissions, credit lending, agricultural lending, a multi-domain benchmark corpus, healthcare, and education. Each is scored on demographic parity and equalized odds difference, plus disparate impact ratio and accuracy cost where computable. Intervention effectiveness tracks baseline disparity magnitude: across model-domain pairs the constraint improved disparity in 9 of 14 high-disparity cases and worsened it in 3 of 4 near-fair ones. Each of the five high-disparity exceptions reverses under one of two measurement checks, a minimum group size or thresholds fit on held-out data. A CUSUM monitor started at deployment, tested on a simulated shift, separates constrained models that never met a 0.1 parity convention from those that met it and later regressed. A single deployment-time audit is therefore an unreliable guide, which argues for baseline-disparity screening and continuous monitoring

著者のコメント

18 pages, 6 figures, 2 tables. Code, data loaders and figures: github.com/nithinnarla/fape-fairness-ml

arXiv ID: 2609.26955 / 要約の誤りについて