強い近視を含む多民族集団の緑内障を眼底写真から検出
Detecting Glaucoma Across Multi-ethnic Myopic and Non-Myopic Populations Using an Uncertainty-Aware Vision Transformer: A Multicentre Model Development and Validation Study
この論文をやさしく読む
ひとことで言うと
眼底写真から緑内障を見つけるモデルを、近視の有無や地域が異なるデータで検証した。
何に役立つ?
高度近視の多い集団での緑内障スクリーニング支援が想定される。臨床導入の効果を実証したという報告ではない。
この研究の面白いところ
56,483枚で開発し、3大陸16データセットで検証した。高度近視の眼と通常の眼を分けた性能も報告している。
どこまで分かった?
外部データのAUROCは86.4~99.6%と幅がある。医師との比較は高度近視の探索的評価で、使用できた情報の範囲も異なる。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
背景:カラー眼底写真によるAIの緑内障検出は大規模なスクリーニングに利用できる可能性があるが、正解の定義、集団、高度近視などの併存状態が違う外部データでは性能が低下し得る。著者らは、高度近視の有無を含む多民族の集団で緑内障を検出するVision Transformerモデルを開発・検証した。方法:56,483枚の眼底写真を用い、予測の不確実性を推定するViT-B/16モデルを開発した。写真の57.1%は近視、14.4%は高度近視で、緑内障のラベルは臨床所見、画像、視野検査のデータを用いて標準化した。高度近視のラベルが明示された4データセットを含む、3大陸の独立した16データセットで検証した。結果:内部検証のAUROCは98.7%(95%信頼区間98.2~99.1%)、感度94.5%、特異度97.3%だった。8か国の16の外部データセットではAUROCは86.4~99.6%だった。高度近視の眼では内部AUROCが97.8%(95%信頼区間96.1~99.2%)、感度94.8%、特異度93.7%だった。外部の高度近視データでのAUROCは北京眼研究で86.5%、台湾、タイ、韓国の病院データでそれぞれ93.3%、91.8%、85.5%だった。探索的な高度近視の臨床評価では、眼底写真だけを用いる診断の正確度は眼科医と訓練を受けた判定者より高く(92.0%対70.0%、p=0.008)、全臨床情報を用いる緑内障専門医と同程度だった。解釈:モデルは近視の有無を含む多民族集団で緑内障検出性能を示し、高度近視が多い環境でAIを用いたスクリーニングを支援する可能性がある。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Background: Artificial intelligence (AI)-based glaucoma detection from colour fundus photographs (CFP) offers scalable screening, but performance may decline on external datasets because of differences in ground-truth definitions, populations, and coexisting conditions such as high myopia (HM). We developed and validated a Vision Transformer-based deep learning (DL) model for glaucoma detection across multi-ethnic cohorts with and without HM. Methods: A ViT-B/16 model with predictive uncertainty estimation was developed using 56,483 CFPs (57.1% with myopia; 14.4% with HM). Glaucoma labels were standardised using clinical, imaging, and perimetry data. The model was validated on 16 independent datasets across three continents, including four datasets with explicit HM labels. Findings: Internal AUROC was 98.7% (95% CI 98.2-99.1%), with sensitivity 94.5% and specificity 97.3%. Across 16 external datasets from eight countries, AUROCs ranged from 86.4% to 99.6%. In HM eyes, internal AUROC was 97.8% (95% CI 96.1-99.2%), with sensitivity 94.8% and specificity 93.7%. External HM AUROCs were 86.5% in the Beijing Eye Study and 93.3%, 91.8%, and 85.5% in hospital-based datasets from Taiwan, Thailand, and South Korea. In an exploratory HM clinical evaluation, the model had higher CFP-only diagnostic accuracy than ophthalmologists and trained graders (92.0% vs 70.0%; p=0.008) and performed comparably to glaucoma specialists using full clinical information. Interpretation: The model showed robust glaucoma detection across myopic and non-myopic multi-ethnic populations and may support AI-assisted screening in settings with high HM prevalence.
arXiv ID: 2609.29433 / 要約の誤りについて