arXiv論文メモ
新着一覧
stat.ME / math.ST / stat.TH · 査読状況未確認

多数の小標本検定で誤発見を抑えつつ検出力を保つ

Empirical Bayes prepivoting under group invariance: false discovery rate control and moderated t-tests

Nikolaos Ignatiadis and Etienne Roquain

この論文をやさしく読む

ひとことで言うと

各項目の測定数が少ない一方で、調べる項目が非常に多い検定を扱います。項目間で分散情報を共有しながら、誤って有意とする発見の割合を制御する統計手法です。

何に役立つ?

多数の遺伝子やタンパク質を同時に調べるような解析に関係します。少標本の検出力と誤発見率の保証を両立させるための理論的な方法を提供します。

この研究の面白いところ

分散の学習に使う情報を、帰無分布を保つ変換の軌道を通じたものに限定します。この構成によって、limmaの階層モデルを仮定しない有限標本の保証と、別の条件下での高い漸近検出力を結び付けます。

どこまで分かった?

有限標本でのFDR制御には単位間の独立性と帰無分布の群不変性が必要です。オラクルと同じ検出力という結論は、limmaの作業モデルと疎な漸近領域に関するもので、任意の実データで同じ性能を保証するものではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

遺伝子やタンパク質など数千の単位について、それぞれ少数の反復測定から平均がゼロかどうかを同時に検定する問題を考える。ゲノミクスで広く使われ、limmaソフトウェアに実装されている方法は、単位間で情報を共有して各単位に固有の分散の分布を学び、その後、分散推定を調整したt統計量を計算する。 本研究では、この調整済みt統計量を用い、次の二つを満たす初の手続きを開発する。第一に、単位間の独立性と帰無分布の群不変性の下で、limmaの階層モデルを仮定せず、有限標本で誤発見率(FDR)を制御する。第二に、limmaの作業モデルに従う疎な漸近領域で、オラクルの局所誤発見率手続きと同じ検出力を達成する。同じ領域では、通常のt検定のp値にBenjamini–Hochberg(BH)法を適用すると、検出力は漸近的にゼロになる。 提案法は、符号反転や直交回転など、帰無分布を保つ変換のコンパクト群を使う。各単位のデータに、その群による軌道だけを通じて依存する統計量から、分散の分布を学習する。次に、得られた調整済みt統計量を、単位をまたいで集めた群変換後の統計量に照らして較正し、複合p値を得る。これをBH法およびその近い変種と組み合わせて使う。小さな有限群については、同じ学習済みスコアを使うSelective SeqStep+手続きも構成する。この方法は、二標本検定と線形モデルの係数の検定にも拡張できる。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

We consider simultaneously testing hypotheses about thousands of units, e.g., genes or proteins, where each unit yields a handful of replicate measurements and we test whether its mean is zero. A widely used approach in genomics, implemented in the limma software, borrows strength across units to learn the distribution of the unit-specific variances, then computes moderated t-statistics. Here we develop the first procedures using moderated t-statistics that (i) control the false discovery rate (FDR) in finite samples under independence across units and null group invariance, without assuming limma's hierarchical model, and (ii) match the power of an oracle local false discovery rate procedure in a sparse asymptotic regime under limma's working model. Benjamini-Hochberg (BH) with standard t-test p-values has asymptotically zero power in the same regime. Our approach learns the variance distribution from statistics that depend on each unit's data only through its orbit under a compact group of transformations that preserves the null distributions, such as sign flips or orthogonal rotations. We then calibrate the resulting moderated t-statistics against group-transformed statistics pooled across units to obtain compound p-values, which we use with BH and a close variant. For small finite groups, we also construct a Selective SeqStep+ procedure using the same learned scores. Our approach extends to two-sample tests and tests of linear model coefficients.

arXiv ID: 2609.26764 / 要約の誤りについて