被覆率が同じでも共分散近似で統計的判断が変わる
Coverage Is Not Ordering: Ancillary Leakage and Representation Dependence in Covariance-Based Inference
この論文をやさしく読む
ひとことで言うと
信頼区間の被覆率が正しくても、共分散による近似で同じデータに対する判断が変わることを示しています。
何に役立つ?
測定モデルを平均と共分散だけで置き換える推論で、被覆率の点検だけでは見逃す影響を確認するために役立ちます。
この研究の面白いところ
本来パラメータの情報を持たない残差が、近似を通じて判断へ混入する機構を解析します。相対不確かさ10%の例で、両方の被覆率を正確に較正しても約7.6%の判断が変わります。
どこまで分かった?
解析可能な相関測定モデルでの理論結果と例です。被覆のずれと判断順序のずれは異なる次数で現れ、被覆率が同じことから推論全体が等価とは結論できないという範囲の主張です。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
共分散行列は、測定の統計モデルをコンパクトに代用するものとしてよく使われる。本研究では、この置き換えが、当てはめた値やその不確かさだけでなく、実験そのものの尤度比による順序付けも変え得ることを示す。扱いやすい相関測定モデルでは、データは情報を持つ成分と補助統計量である残差の不一致に分かれる。したがって厳密な尤度推論では、その不一致を使わない。しかし、データに依存する共分散行列や、非線形変換後のガウス分布による再構成は、それをパラメータ推論へ再び持ち込む可能性がある。 その結果生じる変化を、同じように較正した受容領域の対称差の確率質量で定量化する。これは、実験を繰り返したとき、検定対象のパラメータ値について信頼区間による判断が変わる実験の割合である。不確かさが小さい領域では、順序付けの食い違いは一般に総相対不確かさの1次で現れる一方、通常の被覆率のずれと補助統計量で条件付けた被覆率のずれは2次から始まる。「被覆率では見えない」点では、被覆率の差の主要項が消えるのに、順序付けの食い違いはゼロにならない。総相対不確かさが10%の代表的な例では、両手続きを厳密に較正していても、約7.6%の実験で信頼区間による判断が変わる。 同じ仕組みはPeelleのPertinent Puzzleを、補助統計量と関連部分集合に関する古典的な文献につなぐ。つまり、共分散近似は、真のモデルでは補助統計量である適合度統計量に、推論上の関連性を作り出し得る。同じ情報内容に対して報告される信頼区間が、これによってどう変わるかを直接示し、統計的不確かさが等しい場合や測定が二つの場合を越えて解析を拡張する。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-17(UTC)
- 最新改訂
- 2026-09-17 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-17 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
A covariance matrix is often used as a compact surrogate for the statistical model of a measurement. We show that this replacement can change not only a fitted value or its uncertainty, but also the likelihood-ratio ordering of the experiment itself. In a tractable correlated-measurement model, the data separate into an informative component and an ancillary residual disagreement. Exact likelihood inference therefore does not use that disagreement, whereas data-dependent covariance matrices and Gaussian reconstructions after nonlinear transformations can reintroduce it into parameter inference. We quantify the resulting change by the probability mass of the symmetric difference of equally calibrated acceptance regions - the fraction of repeated experiments for which the confidence decision about a tested parameter value changes. In the small-uncertainty regime the ordering discrepancy is generically first order in the total relative uncertainty, whereas conventional and ancillary-conditioned coverage defects begin at second order. At 'coverage-blind' points the leading coverage difference vanishes while the ordering discrepancy remains nonzero. In a representative blind case with 10% total relative uncertainty, about 7.6% of experiments change their confidence decision despite exact calibration of both procedures. The same mechanism connects Peelle's Pertinent Puzzle to the classical literature on ancillarity and relevant subsets: a covariance approximation can manufacture inferential relevance for a goodness-of-fit statistic that is ancillary in the true model. We show directly how this changes the confidence interval reported for the same informative content, and extend the analysis beyond equal statistical uncertainties and beyond the two-measurement case.
著者のコメント
34 pages, 8 figures
arXiv ID: 2609.20007 / 要約の誤りについて