疾病診断の分類表で関連を図示・検定する二段階法
A two-step log-linear procedure for graphical representation and inference of associations in cross-classified data for disease diagnosis
この論文をやさしく読む
ひとことで言うと
疾病診断などの分類データで、項目間の関連を図に表し、差の原因を検定する統計手順。
何に役立つ?
二値の属性を組み合わせた診断データの関連を解釈するために使える。要旨では冠動脈疾患の実例に適用した。
この研究の面白いところ
図の中の位置の一致とオッズ比の組の一致が、検定として同じになることを示した。
どこまで分かった?
図による解釈は要旨で述べるカテゴリ数やデータ行列の条件に基づく。診断精度そのものの改善を実証したとは述べていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
生物統計、特に疾病診断では、複数の分類を組み合わせたデータの関連を解析することが多い。距離による関連モデルは、カテゴリ数が少なく、疎でない行列について図による解釈を与える。この枠組みでは探索変数と応答変数が二値であることが多く、個々の属性の組合せに基づく解析が重要になる。本研究は、飽和モデルについて、距離による関連のパラメータ化でも、対数線形モデルの通常の線形関係が全次元で保たれると示す。これにより、全体効果と主効果を除いた後に、展開図で関連を解析・解釈する二段階の手順を作れる。提案手順は二値変数で表された属性の組合せを持つ分類表を扱え、従来の統計ソフトで実装しやすい。 疾病診断に関しては、展開図で解が退化する問題と、属性の組合せの位置に有意な差があるか判定する問題に対処する。オッズ比に基づく独立性の仮説検定を考え、検定が有意になる原因を誤りの伝播を避けて特定する手順も提案する。疾病診断の展開図において、二つの属性の組合せの位置が等しいという検定と、対応するオッズ比の組が等しいという検定が同値であることを示す。結果を冠動脈疾患の診断に関する実例へ適用し、オッズ比と診断検査の性能指標を関連付けた。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-23(UTC)
- 最新改訂
- 2026-09-23 · v1
- 査読・掲載
- 掲載先の記載あり
著者による掲載先の記載:Statistics in Medicine, 2023, 42, 4897-4916。出版社での独立確認は未実施です。
更新履歴
- v1 2026-09-23 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Biometrical sciences and disease diagnosis in particular, are often concerned with the analysis of associations for cross-classified data, for which distance association models give us a graphical interpretation for non-sparse matrices with a low number of categories. In this framework, usually binary exploratory and response variables are present, with analysis based on individual profiles being of great interest. For saturated models, we show the usual linear relationship for log-linear models is preserved in full dimension for the distance association parameterization. This enables a two-step procedure to facilitate the analysis and the interpretation of associations in terms of unfolding after the overall and main effects are removed. The proposed procedure can deal with cross-classified data for profiles by binary variables, and it is easy to implement using traditional statistical software. For disease diagnosis, the problems of a degenerate solution in the unfolding representation, and that of determining significant differences between the profile locations are addressed. A hypothesis test of independence based on odds ratio is considered. Furthermore, a procedure is proposed to determine the causes of the significance of the test, avoiding the problem of error propagation. The equivalence between a test for equality of odds ratio pairs and the test for equality of location for two profiles in the unfolding representation in the disease diagnosis is shown. The results have been applied to a real example on the diagnosis of coronary disease, relating the odds ratios with performance parameters of the diagnostic test.
著者のコメント
20 pages, 6 tables, 1 figures
arXiv ID: 2609.27491 / 要約の誤りについて