舌と脈の画像を統合して中医学の診断を行うLingLan
LingLan: An Advancing Traditional Chinese Medicine Diagnosis LLM with Multimodal Data
この論文をやさしく読む
ひとことで言うと
舌や脈の画像を構造化して中医学の四診の情報と統合し、専用の言語モデルで診断を行う研究。
何に役立つ?
中医学の診断データをデジタル化し、複数の診断情報を扱うモデルを評価する際の参考になる。
この研究の面白いところ
基準手法に対し報告された正確さは30.82%から62.72%へ上がった。
どこまで分かった?
示された数値は研究内のモデル評価結果であり、患者での臨床的な安全性や治療効果を実証したとは要旨にない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
AIは現代医学を変えつつあるが、中医学への導入は比較的遅い。中医学が、望診、聞診、問診、切診という全体的で主観的な診断方法に依存し、定量的で標準化された医療体系と合わせにくいことが主な理由である。本研究は、舌と脈の画像を構造化された臨床上の標準的な記述へ自動的に変換し、複数の情報源を四診の過程に沿った一つのデジタル記録へ統合するMultimodal Data統合枠組み(UFMD)を導入する。この構造化データを基に、四診の診断論理と手順を模倣するよう教師あり学習で追加学習した、中医学専用の言語モデルLingLan-14Bを作る。実験では診断の正確さが基準の30.82%から62.72%へ上がり、相対改善率は103.5%だった。F1値は最大82%に達した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-22(UTC)
- 最新改訂
- 2026-09-22 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-22 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Though artificial intelligence (AI) increasingly transforms modern medicine, its integration into Traditional Chinese Medicine (TCM) has been relatively slow, primarily due to TCM's reliance on holistic, subjective diagnostic methods---namely Inspection, Auscultation and Olfaction, Inquiry, and Palpation(I-AOI-P)---which are difficult to align with quantitative, standardized medical systems. In this work, we introduce a Unification Framework for Multimodal Data (UFMD), which automatically processes tongue and pulse images into structured, clinically standard descriptions, integrating multi-source diagnostic information into a unified digital record of I-AOI-P process. Building on this structured data, we create LingLan-14B, a TCM-specific large language model fine-tuned via supervised learning to emulate the diagnostic logic and workflow of I-AOI-P process. Experimental results show that our method significantly enhances diagnostic accuracy, achieving a relative improvement of 103.5% over the baseline (62.72% vs. 30.82%) and reaching an F1-score of up to 82%.
著者のコメント
6 pages, 5 figures
arXiv ID: 2609.25715 / 要約の誤りについて