arXiv論文メモ
新着一覧
cs.AI · 査読状況未確認

診断の失敗から医療AIの機能を増やすMedRSI

MedRSI: Recursive Self-Improvement for Medical Agents via Clinically Aligned Self-Evolution

Junde Wu, Jiayuan Zhu, Minghao Hu, Fenglin Liu, Jiazhen Pan

この論文をやさしく読む

ひとことで言うと

医療AIが診断で失敗したときに、その原因に対応するツールやモデルを作り、後の患者データでも役立つものを残していく仕組みです。

何に役立つ?

開発者が最初に用意しなかった機能を、診断上の課題に応じて追加する方法として検討できます。公開ベンチマークと非公開タスクでの結果が示されていますが、医療現場での無監督運用を保証するものではありません。

この研究の面白いところ

失敗の回数だけでなく臨床上の影響を優先順位に使い、機能を作る速さと正式に採用する慎重さを分けています。新しい機能を増やし続ける際の採用基準が研究の中心です。

どこまで分かった?

初の医療向け枠組みという位置付けは著者の主張です。要旨には改善の具体的数値や患者数、前向き試験による安全性評価は示されていません。ベンチマーク性能と患者の転帰改善は区別する必要があります。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

医療エージェントでは、汎用的な推論モデルと専門的な臨床ツールの組み合わせが増えている。しかし、その能力は依然として、導入前に臨床家や技術者が設計した範囲に大きく固定されている。再帰的自己改善(RSI)は、エージェントが自らの失敗から学び、自律的に能力を拡張する別の枠組みを提供するが、医療にそのまま適用すると根本的な安全性の課題が生じる。 本研究では、ツールの組み合わせとタスクに特化したモデル学習を通じて、診断の失敗を新たな臨床能力へ継続的に変換する、医療向けとして初の再帰的自己改善フレームワークMedRSIを提案する。臨床実践に着想を得て、臨床的な要請に沿う自己進化のための2つの仕組みを導入する。臨床的コストを考慮した失敗の優先順位付けは、頻度だけでなく、起こり得る臨床上の影響に応じて改善対象を選ぶ。「素早い発見と慎重な登録」は、迅速な機能の考案と保守的な採用を分離し、その後の患者集団でも継続的な利点を示した新しいツールだけを、持続的に使うエージェントへ取り込む。 公開の緑内障・心疾患ベンチマークと2つの非公開の臨床タスクにおいて、MedRSIは領域分割、計測、予測、マルチモーダル推論、生成の能力を段階的に開発した。手作業で設計された医療エージェントを上回り、当初の設計者が想定していなかった臨床上の問題への解決策を自律的に発見した。これらの結果は、医療エージェントの能力が導入前の仕様に縛られ続ける必要はなく、何を改善し何を残すかを臨床的な根拠に基づいて管理すれば、診断経験から新しい能力を継続的に構築、検証、蓄積できることを示す。コードはhttps://github.com/ImprintLab/MedRSIで公開されている。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Medical agents increasingly combine general reasoning models with specialized clinical tools, yet their capabilities remain largely fixed by what clinicians and engineers design before deployment. Recursive self-improvement (RSI) offers a different paradigm in which agents learn from their own failures and autonomously expand their capabilities, but directly applying RSI to medicine introduces fundamental safety challenges. We introduce MedRSI, the first recursive self-improvement framework for medicine, which continuously transforms diagnostic failures into new clinical capabilities through tool composition and task-specific model training. Inspired by clinical practice, MedRSI introduces two mechanisms for clinically aligned self-evolution. Clinical-cost-aware failure prioritization directs improvement toward errors according to their potential clinical consequences rather than frequency alone. Fast discovery with slow registration separates rapid capability invention from conservative adoption, allowing new tools to enter the persistent agent only after demonstrating sustained benefit across subsequent patient cohorts. Across public glaucoma and heart disease benchmarks and two private clinical tasks, MedRSI progressively develops segmentation, measurement, prediction, multimodal reasoning, and generative capabilities, surpasses manually engineered medical agents, and autonomously discovers solutions to clinical problems not anticipated by its original designers. Our results show that medical agents need not remain constrained by capabilities specified before deployment: with clinically grounded mechanisms governing what to improve and what to retain, they can continuously construct, validate, and accumulate new capabilities from diagnostic experience. Code is available at https://github.com/ImprintLab/MedRSI.

arXiv ID: 2609.24838 / 要約の誤りについて