医用画像の長い裾の分布に合わせて注釈する標本を選ぶ
Less Is More in the Long Tail: Stage-Adaptive Sample Selection for Annotation-Efficient Dense Prediction
この論文をやさしく読む
ひとことで言うと
医用画像の注釈予算を、少ないカテゴリや学習段階を考慮して配分し、限られた注釈で分割精度を保つ方法。
何に役立つ?
画素やボクセル単位の注釈が高価な画像分割で、どの標本に注釈するかを決める参考になる。
この研究の面白いところ
40%の注釈予算で全データ学習の98.3%の性能を得て、特定の難しい構造では全データ学習を上回った。
どこまで分かった?
結果は108構造を含む医用画像分割の評価環境でのもの。98.3%は全データ性能に対する比率で、絶対的な診断精度を示さない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
深層学習は一般に学習データが増えると性能が向上するが、カテゴリの頻度が偏る大規模な密な予測課題では、画素やボクセルごとの注釈に多大な費用がかかり、その拡大が制約される。本研究は、プール型能動学習で学習段階に応じてデータを選ぶSASSを提案する。ラベルなしで使える自己教師ありの勾配スコア、事前情報と検証結果のフィードバックに基づくカテゴリの再均衡、モデルの学習過程に合わせた取得方法を組み合わせる。勾配スコアの計算に候補画像の正解マスクは使わず、頻度の低いカテゴリや表現の変化に応じて選び方を変える。108の解剖学的構造にまたがる10万件超の多モーダルな三次元医用画像分割の評価では、全データでの性能の98.3%を、学習プールの注釈予算40%で再現し、BADGEを5.1パーセントポイント上回った。さらに、注釈量を減らした方が良いという統計的に裏付けられた傾向があり、難しい構造の群全体、ならびに膵臓と胆嚢の構造別評価で、全データを使う学習を上回った。結果は、注釈効率が選ぶ標本だけでなく、カテゴリ間の予算配分と、モデル由来の点数を選択に使い始める時期にも左右されることを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-22(UTC)
- 最新改訂
- 2026-09-22 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-22 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Deep learning performance generally improves with increasing training data, yet this scaling is fundamentally constrained by annotation cost in large-scale dense prediction tasks with long-tailed category distributions, where pixel- or voxel-level annotation is prohibitively expensive. We propose SASS (Stage-Adaptive Sample Selection), a stage-adaptive data-selection framework for pool-based active learning in long-tailed dense prediction. SASS combines three components: label-free self-supervised gradient scoring, prior-guided category rebalancing with validation-driven feedback, and stage-adaptive acquisition aligned with model training dynamics. This design avoids candidate ground-truth masks during gradient scoring while making acquisition responsive to long-tail imbalance and evolving representations. We evaluate SASS on a multimodal 3D medical segmentation testbed comprising over 100,000 samples spanning 108 anatomical structures. SASS recovers 98.3% of full-dataset performance with a 40% training-pool annotation budget, outperforming BADGE by 5.1 percentage points. Moreover, SASS exhibits a statistically supported less-is-more pattern, surpassing full-dataset training at the Hard-group level and, at the structure level, for the pancreas and gallbladder. More broadly, SASS shows that annotation-efficient learning depends not only on which samples are selected, but also on how the annotation budget is distributed across categories and when model-derived scores begin to guide selection.
arXiv ID: 2609.25850 / 要約の誤りについて