追跡期間を説明変数に混ぜたCox回帰のずれを解析
Least-false Cox coefficients under affine follow-up contamination: exact continuous- and grouped-time benchmarks
この論文をやさしく読む
ひとことで言うと
追跡後に分かる情報を開始時点の説明変数として扱うと、Cox回帰の係数と不確実性がどうずれるかを解析した研究です。
何に役立つ?
生存時間解析で追跡期間から作った変数を使う際、推定対象や標準誤差への影響を理解する助けになります。
この研究の面白いところ
連続時間では縮小と飽和が見られる一方、打ち切りや時刻のグループ化によって係数の振る舞いが大きく変わります。
どこまで分かった?
指定したアフィンな汚染クラスと条件の下での解析です。サンドイッチ推定は不確実性を推定しても、係数の推定対象自体は修正しません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
対象者の追跡が終わった後に集計した共変量を、あたかも開始時点で観測したかのようにCox回帰へ入れることがある。この方法は将来のイベントや打ち切りの情報を取り込み、推定対象と標本上の振る舞いの両方を変える。本研究は、真の開始時点の共変量が実際の追跡時間によって汚染されるアフィンなクラスを解析する。部分尤度を観測データに基づく推定基準として扱い、母集団スコアを導出し、その一意な「最も誤りの少ない」係数を特徴づける。 厳密な連続時間の基準計算では、汚染が強いと尺度の縮小と飽和が起こる。一方、管理上の打ち切りはこの縮小を崩し、過大な値を生み得る。退出時刻をBreslowの同時刻処理でグループ化すると幾何が変わり、誘導された関連が発散しても、係数はいったん高くなった後、最終的にゼロへ戻る。さらに観測データの影響関数を導き、モデルに基づく分散が過小にも過大にもなり得る理由を示す。対象者単位のサンドイッチ推定は、記載した条件の下では、この最も誤りの少ない推定対象の周囲の不確実性を一貫して推定するが、推定対象そのもののずれは直さない。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Covariates summarized over a subject's completed follow-up are sometimes entered into Cox regression as though observed at baseline. This practice incorporates future event or censoring information and changes both the estimand and its sampling behavior. We analyze an affine class in which a genuine baseline covariate is contaminated by realized follow-up time. Treating partial likelihood as an observed-data estimation criterion, we derive the population score and characterize its unique least-false coefficient. Exact continuous-time benchmarks show scale reduction and saturation under strong contamination, while administrative censoring destroys the reduction and may produce overshoot. Grouping exit times with Breslow ties changes the geometry: the coefficient has a single hump and eventually returns to zero even though the induced association diverges. We also derive an observed-data influence function and show why model-based variance can be either too small or too large. A subject-level sandwich consistently estimates uncertainty around the least-false target under the stated conditions, but it does not correct the target itself.
arXiv ID: 2609.28872 / 要約の誤りについて