arXiv論文メモ
新着一覧
cs.CV · 査読状況未確認

少数の病変画像からカプセル内視鏡画像を生成するEndoFSA

EndoFSA: Endoscopic Few-Shot Image Generation via Rank-Constrained Parameter Adaptation

Panagiota Gatoula, Grigoris Karypidis, Dimitris K. Iakovidis

この論文をやさしく読む

ひとことで言うと

少数しかない病変のカプセル内視鏡画像を補うため、正常画像で学んだ生成器を小さく調整して合成画像を作る研究です。

何に役立つ?

異常画像が少ない場合の分類器の学習データを補う用途が考えられます。要旨では合成異常画像だけで学習した分類器が、実画像を使った場合と同程度の性能だったと報告しています。

この研究の面白いところ

事前学習済みの重みを固定して少数の調整パラメータだけを更新し、画素単位の注釈や病変マスクなしで適応します。

どこまで分かった?

評価は公開WCEベンチマークの画像生成と分類課題についてです。実際の臨床診断での安全性や有効性は要旨に示されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

カプセル内視鏡(WCE)では大量の消化管画像が得られるが、病的所見の画像は著しく少なく、深層学習による異常検出システムの汎化性能を制限する。合成データ生成はこの偏りを緩和する実用的な方法だが、少数の異常画像で直接訓練すると、不安定さ、過学習、構造の歪みが生じやすい。解剖学的な事前知識を保ちながら現実的な病変の変化を表せる、制御された適応方法が必要である。本研究は、WCEで少数の画像から内視鏡画像を生成する、GANに基づくモデルEndoFSAを提示する。豊富な正常画像で事前学習した生成器を使い、限られた学習例で異常領域に適応させる。この際、事前学習済みの重みは固定し、少数の調整パラメータだけを更新する、ランク制約付きのパラメータ適応を用いる。 更新を低次元部分空間に制限し、知覚的な境界正則化とクラスタごとの多様性制御を組み込むことで、少ないデータでも効率よく適応し、正常画像から学んだ解剖学的な事前知識を保ちながら、生成画像が似たものに偏るモード崩壊を緩和する。画素単位の注釈、マスク、境界枠の教師信号は不要である。さまざまな異常の種類を含む公開WCEベンチマークで、EndoFSAは実際の病変形態を再現する異常画像を生成した。さらに、EndoFSAが作った合成異常画像だけで画像分類器を訓練した後段の分類課題では、実画像で訓練した場合と同程度の性能が得られた。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

WCE produces large-scale gastrointestinal image data yet pathological findings remain significantly underrepresented limiting the generalization performance of deep-learning based abnormality detection systems. SDG methods offer a practical solution to mitigate this imbalance. However their training directly on scarce abnormal samples often results in instability overfitting and structural distortions. Addressing these challenges requires controlled adaptation mechanisms that preserve anatomical priors while enabling realistic pathological variation. This paper presents EndoFSA a GAN-based model for Endoscopic Few-Shot image generation by Adaptation in WCE imaging. EndoFSA leverages a generator pretrained on abundant normal data and adapts it to abnormal domains using limited number of training samples through a rank-constrained parameter adaptation where only a small number of modulation parameters is updated while the pretrained weights remain frozen. By restricting parameter updates to a low dimensional subspace and incorporating perceptual boundary regularization and cluster-wise diversity control EndoFSA enables efficient model adaptation under limited data conditions and mitigates mode collapse while preserving the anatomical priors learned from normal data. Importantly EndoFSA operates without requiring pixel-level annotations, masks or bounding box supervision. Evaluation on publicly available WCE benchmark datasets spanning various abnormal categories demonstrates that EndoFSA generates abnormal images reproducing real lesions morphology. Moreover in a downstream classification task training an image classifier solely on synthetic abnormal images generated by EndoFSA yields performance comparable to that obtained with real images.

著者のコメント

Presented at the 39th IEEE International Symposium on Computer-Based Medical Systems (CBMS 2026), June 2026

arXiv ID: 2609.29930 / 要約の誤りについて