arXiv論文メモ
新着一覧
stat.ML / cs.LG · 査読状況未確認

BTL選好モデルの最尤推定量が成立する標本数と誤差

Error Bounds for Statistical Estimators in BTL Model with Parametric Multivariate Utility Functions

Yicheng Li and Huifu Xu

この論文をやさしく読む

ひとことで言うと

一対比較から好みの強さを推定するBTLモデルで、制約のない最尤推定が成立する標本数と誤差を調べた研究。

何に役立つ?

選好調査でどの組を何回比較するか設計し、推定結果の信頼性を理論的に評価する際の参考になる。

この研究の面白いところ

一様でない決定的な質問設計を許し、標本数のしきい値、誤差の内訳、ミニマックス速度を同じ枠組みで扱う。

どこまで分かった?

理論は同時識別可能性や有界なダイナミックレンジなどの条件を要する。数値結果は要旨では予備的とされ、具体的な誤差値は示されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

Bradley–Terry–Luce(BTL)モデルの下で選好を引き出す問題を研究する。真の部分価値ベクトルは未知であり、収集した選好情報からパラメータとして推定する必要がある。選ばれる一対比較の質問集合は一様でなく、決定的で、選択肢全体に対して任意でよいが、同時識別可能性の条件を満たすものとする。特に、実行可能集合へ明示的な有界性制約や外部の正則化を加えずに、標準的な最尤推定量が有限となり、鋭い誤差上界を持つ条件を調べる。 そのために、標準的な有界なダイナミックレンジ条件の下でミニマックス下界を導き、古典的なCramér–Rao下界に現れるのと同じFisher情報の幾何が、有限標本での推定の難しさの基礎にあることを見いだす。尤度のスコア方程式の非漸近的な展開と固定点による局所化を組み合わせ、ある設計に依存する標本数のしきい値を超えると、制約なしの標準的な最尤推定量が高確率で存在し一意になることを示す。同じ展開から、推定誤差を線形な確率項、明示的な二次のバイアス、高次の剰余に分解する。さらに詳しい解析で、標準的な最尤推定量が対数因子と定数因子を除いてミニマックス速度を達成するための十分な標本数条件を与える。これにより、パラメータを持つ効用の推定について統一的な非漸近理論を示し、外部の正則化でなく回答データだけから推論が決まる条件を明らかにする。予備的な数値結果は理論と整合する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

We study preference elicitation under the Bradley-Terry-Luce (BTL) model where the true partworth vector is unknown and has to be estimated as a parameter with elicited preference information. The set of selected pairwise queries is non-uniform, deterministic, and arbitrary over a collection of alternatives, provided that it satisfies a joint identifiability condition. We focus on understanding when the canonical maximum likelihood estimator (MLE) is finite and admits sharp error bounds without explicit compactness constraints on the feasible set or external regularizers. To this end, we derive minimax lower bounds under the standard bounded dynamic range condition, and find that the same Fisher-information geometry in the classic Cramér-Rao lower bounds underpins the finite-sample difficulty of the estimation problem. By combining a non-asymptotic expansion of the likelihood score equation with a fixed-point localization argument, we identify a design-dependent sample size threshold above which the unconstrained canonical MLE exists and is unique with high probability. The same expansion yields a decomposition of the estimation error into a linear stochastic term, an explicit second-order bias, and a higher-order remainder. A refined analysis gives sufficient sample size conditions under which the canonical MLE attains the minimax rates up to logarithmic and constant factors. These results provide a unified non-asymptotic theory for parametric utility elicitation and reveal when the inference is determined by response data alone rather than by external regularization. Preliminary numerical results are consistent with the theoretical findings.

arXiv ID: 2609.26326 / 要約の誤りについて