arXiv論文メモ
新着一覧
cs.LG / cs.AI · 査読状況未確認

医療AIがリスクと誤判定費用を判断に統合できるか検証

Misaligned Clinical Risk Classification and Cost Asymmetry in Open-Weight Large Language Models

Star S.D. Liu, Xiyu Ding, Robert B. Barrett, Alberto Santamaria-Pang, Nic Dobbins, Harold P. Lehmann

この論文をやさしく読む

ひとことで言うと

病気を見逃す損失と、病気でない人を誤って陽性にする損失を伝えたとき、AIの判断が適切な方向へ変わるかを調べています。内部にリスクや費用の情報があっても、最終判断に正しく使えるとは限りませんでした。

何に役立つ?

医療AIの評価で、予測精度だけでなく誤判定の重みを変えたときの応答も点検する根拠になります。患者への診療方針を示す研究ではなく、公開データでモデルの振る舞いを評価したものです。

この研究の面白いところ

AUC約0.83でリスク情報を内部から読み出せても、費用に沿った二方向の応答と順序をともに満たすのは12条件中2条件でした。情報を表現する能力と、判断に統合する能力を分けて調べています。

どこまで分かった?

四つの公開重みモデル、糖尿病の公開データ、11費用比、三つの言い回しでの評価です。AUCは内部リスクの線形読み出しに関する値で、全条件で費用を考慮した臨床判断が正しかったという意味ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

大規模言語モデル(LLM)が、患者のリスクと臨床上の費用のトレードオフをどのように統合するかは、十分に理解されていない。本研究では、重みが公開された四つのLLM、Qwen-2.5-7B/32BとLlama-3.1-8B/70Bについて、費用のトレードオフを内部でどう表現するか、その表現が臨床予測とどう関係するか、指定した費用の方向と大きさから予想されるとおりに判断が変わるかを調べた。公開糖尿病データセットを用い、偽陰性(FN)と偽陽性(FP)の費用比を11通りに変え、三つの表現の仕方で内部表現と出力行動を調べた。 患者リスクは従来の分類器に並ぶ水準(AUC ≈ 0.83)で内部表現から線形に読み出すことができ、費用の方向もすべてのモデルから読み出せた。しかし、費用の方向に関する表現の変化が出力の変化に対応したのは、大きい方の二モデルだけだった。また、費用の大きさに対する反応は、主として方向を区別しないものだった。モデルと表現の仕方を組み合わせた12通りのうち、FN費用の増加とFP費用の増加に対して互いに逆向きに反応し、かつ費用に照らして正しい順序を示したのは2通りだけだった。内部表現の面では、一方の費用側で当てはめた方向を他方へ移しても、鏡映対称的な符号化で期待される反転は生じなかった。 これらの結果は、LLMがリスク情報と費用情報を符号化していても、それらを費用に照らして正しい意思決定へ確実に統合するわけではないことを示唆する。したがって臨床評価には、予測性能だけでなく、トレードオフの試験、表現の仕方への感度、既定の動作点も含めるべきである。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

How large language models (LLMs) integrate patient risk with clinical cost tradeoffs remains poorly understood. We investigated how four open-weight LLMs (Qwen-2.5-7B/32B and Llama-3.1-8B/70B) internally represent cost tradeoffs, how these representations relate to clinical predictions, and whether decisions shift as predicted by the specified cost direction and magnitude. Using a public diabetes dataset, we varied 11 false-negative (FN) to false-positive (FP) cost ratios across three phrasings and examined representations and behavioral outputs. Patient risk was linearly recoverable on par with conventional classifiers (AUC $\approx 0.83$), and cost direction was recoverable in every model. However, representational shifts in cost direction tracked output changes only in the two larger models, and responses to cost magnitude were predominantly direction-agnostic. Only 2 of 12 model-phrasings showed both opposing responses to increasing FN versus FP costs and cost-correct ordering. Representationally, a direction fitted on one cost side did not invert when transferred to the other, as expected under mirror-symmetric encoding. These findings suggest that LLMs encode risk and cost information but do not reliably integrate them into cost-correct decisions. Clinical evaluations should therefore include tradeoff tests, phrasing sensitivity, and default operating points alongside predictive performance.

著者のコメント

Submitted to ML4H 2026

arXiv ID: 2609.23999 / 要約の誤りについて