arXiv論文メモ
新着一覧
stat.ME · 査読状況未確認

外部対照データを補正して生存時間解析に生かす

RECaST-Surv: A Calibrated Borrowing Method for Survival Endpoints in Unequal Randomized Trials

Dehua Bi, Arlina Shen, Ruben P.A. van Eijk, Lu Tian, Jiapeng Xu, Guillemette de la Borderie, Nate Bennet, Sarno Maria, Ying Lu

この論文をやさしく読む

ひとことで言うと

対照患者が少ない試験で外部データを使う際に、集団の違いを補正し、誤検出を抑えながら検出力を高める統計手法です。

何に役立つ?

小児疾患や希少疾患など、同時期の対照情報を十分に集めにくい試験の解析が想定されます。報告された改善はシミュレーションと試験を模した評価での結果です。

この研究の面白いところ

外部モデルをそのまま使わず現在の対照群に較正し、さらに検定規則もブートストラップで調整する二段階の工夫があります。

どこまで分かった?

検出力向上は治療効果そのものの増加ではありません。要旨には評価条件ごとの詳細や、実際の新規試験での検証結果は示されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

不均等な無作為割付を含む、同時期の対照群の情報が限られる無作為化試験は、とりわけ小児疾患や希少疾患で倫理的・実務的な利点を持ち得る。しかし、直接比較できる対照患者が少ないため、検出力が低下しやすい。外部対照から情報を借用すれば効率を改善できる可能性がある一方、外部集団と試験集団の比較可能性が十分でないと、第一種過誤率を高めるおそれがある。本研究では、RECaST法を生存時間の設定へ拡張した、イベント発生までの時間を対象とするベイズ転移学習の枠組みRECaST-Survを提案する。 この方法は、外部対照データから構造的生存モデルを学習し、Cauchyランダム効果を介して現在の試験の同時期対照群に較正する。頻度論的な動作特性を改善するため、第一種過誤を制御するよう検定規則を較正する、ブートストラップに基づく手続きも開発する。RECaST-Survは複数の外部データセットを扱え、外部情報源から必要とするのは要約レベルの情報のみである。 シミュレーション研究では、難しい条件にわたって第一種過誤を名目水準付近に保ちながら、標準的な解析より検出力を約10〜12%改善した。筋萎縮性側索硬化症の試験を模した評価では、許容可能な過誤制御を維持しつつ、情報借用を行わない無作為化比較試験解析に対して、検出力を82.8%から95.7%に向上させた。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-16(UTC)
最新改訂
2026-09-16 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Randomized trials with limited concurrent control information---including but not limited to unequal-randomization settings---can offer ethical and practical advantages, especially in pediatric and rare diseases, but they often lose power because fewer control patients are available for direct comparison. Borrowing information from external controls may improve efficiency, but can also inflate the Type I error rate when the external and trial populations are not sufficiently comparable. We propose RECaST-Surv, a Bayesian transfer-learning framework for time-to-event outcomes that extends the RECaST method to survival settings. The method learns a structural survival model from external control data and calibrates it to the concurrent control arm of the current trial through a Cauchy random effect. To improve frequentist operating characteristics, we further develop a bootstrap-based procedure to calibrate the testing rule for Type I error control. RECaST-Surv can accommodate multiple external datasets and requires only summary-level information from external sources. Simulation studies show that the method maintains near-nominal Type I error across challenging settings while improving power by roughly 10%--12% over standard analyses. In an Amyotrophic Lateral Sclerosis trial emulation, RECaST-Surv increased power from 82.8% to 95.7% relative to the no-borrowing RCT analysis, while maintaining acceptable error control.

arXiv ID: 2609.19109 / 要約の誤りについて