arXiv論文メモ
新着一覧
stat.ME · 査読状況未確認

ネットワーク干渉下で偏りなく処置効果を推定

Unbiased Treatment Effect Estimation under Network Interference via Neighborhood-Excluded Cross-Fitting

Haoyang Yu, Anqi Zhao, and Hanzhong Liu

この論文をやさしく読む

ひとことで言うと

人同士のつながりで処置の影響が伝わる実験について、偏りの少ない効果推定の方法を作った。

何に役立つ?

社会ネットワークなど、参加者間の影響を無視できないランダム化実験の分析に役立つ。

この研究の面白いところ

予測モデルの学習時に、評価対象へ影響しうる近傍を除くことで、通常のクロスフィッティングに生じる依存を断つ。

どこまで分かった?

不偏性や推論の保証には指定されたランダム化方式と干渉・分割の条件がある。実験例での短い信頼区間が他の全ネットワークにも当てはまるとは限らない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

個体間に干渉がない場合、クロスフィッティングは柔軟な共変量調整を可能にし、個体ごとの独立なランダム化のもとで有限標本の不偏性も保つ。しかしネットワーク干渉があると、標本外での予測だけでは不偏性が保証されない。評価用分割のHorvitz–Thompson重みに入る割当が、学習標本の結果にも影響し、当てはめた予測と重みの間に依存を生むためである。本研究は、近傍を除外したクロスフィッティングを開発する。推定対象と実験計画に合わせた学習標本を構成し、結果モデルが正しく指定されていなくても、有限標本の不偏性に必要な条件付き独立性を回復する。 ベルヌーイ・ランダム化での直接効果と間接効果、およびベルヌーイ・クラスターランダム化での全体の平均処置効果について、漸近的に妥当な実験計画に基づくWald型推論を確立する。近傍除外により分割数の選択には折り合いが生じる。個体単位の分割では、平均的な除外近傍の大きさとともに分割数を増やす必要がありうるが、クラスター単位ならこの要件を大きく緩め、部分的干渉のもとで固定の分割数を使える。 ベルヌーイ・ランダム化での線形調整については、分散を最小にする方法と信頼区間長を最小にする方法を導き、共変量の次元が増大してよい明示的な速度条件を確立し、分散最適の方法が漸近的に精度を悪化させないことを示す。シミュレーションでは近傍除外を省いたときの偏りと、調整による精度向上を示した。社会ネットワーク実験への適用では、無調整推定量よりもかなり短い信頼区間が得られた。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-20(UTC)
最新改訂
2026-09-20 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Without interference, cross-fitting enables flexible covariate adjustment while preserving finite-sample unbiasedness under independent unit-level randomization. Under network interference, out-of-sample prediction alone no longer guarantees unbiasedness: assignments entering evaluation-fold Horvitz--Thompson weights may also affect outcomes in the training sample, inducing dependence between fitted predictions and those weights. We develop neighborhood-excluded cross-fitting, which constructs estimand- and design-specific training samples to restore the conditional independence needed for finite-sample unbiasedness without a correctly specified outcome model. We establish asymptotically valid design-based Wald inference for direct and indirect effects under Bernoulli randomization and for the global average treatment effect under Bernoulli cluster randomization. Neighborhood exclusion creates a trade-off in choosing the number of folds: unit-level splitting may require the number of folds to grow with average exclusion-neighborhood size, while cluster-level splitting can substantially relaxes this requirement, permitting a fixed number of folds under partial interference. For linear adjustment under Bernoulli randomization, we derive variance-optimal and confidence-interval-length-optimal procedures, establish explicit rate conditions allowing the covariate dimension to diverge, and show that the variance-optimal procedure is asymptotically no-harm. Simulations illustrate the bias from omitting neighborhood exclusion and the precision gains from adjustment. An application to a social network experiment yields confidence intervals substantially shorter than those from the unadjusted estimator.

arXiv ID: 2609.23284 / 要約の誤りについて