画像改変の位置検出を条件別に評価するGIFTBench
GIFTBench: Diagnosing Generalization in Image Forgery Localization and Informing Model Design
この論文をやさしく読む
ひとことで言うと
画像のどこが改変されたかを検出する技術を、改変の種類や対象を分けて評価するデータセットです。診断で見えた問題を基に、検出モデルも設計しています。
何に役立つ?
画像改変検出モデルが、学習時と違う編集方法や画像内容に対応できるかを調べる用途に役立ちます。評価だけでなく、このデータでの学習が外部データへの転移性能を改善したと報告しています。
この研究の面白いところ
115,013枚を複数の軸で整理し、ひとつの平均値では見えない失敗の違いを調べられます。その分析をForenScopeの設計へ結び付けている点も特徴です。
どこまで分かった?
要旨には改善量や各外部データセットの詳細は示されていません。画像単位の改変検出と画素単位の位置特定は別の評価であり、どちらの能力を述べているかを分けて読む必要があります。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
画像改変箇所の位置特定(IFL)を信頼できる形で評価するには、多様な分布変化の下でモデルを評価する必要がある。しかし、既存のベンチマークは改変条件が限られていたり、データセットをまたぐ評価で複数の要因が混ざっていたりすることが多い。その結果、集約した性能だけでは、位置特定の汎化能力を十分に把握できない。 本研究では、画素単位の注釈を持つ115,013枚の改変画像からなる、複数軸のベンチマークGIFTBenchを導入する。改変の生成元、意味的な対象、編集操作、合成の複雑さを網羅し、軸ごとの転移分析と12の外部データセットでの評価を可能にする。診断的な検討から、生成元をまたぐ転移の非対称性、再現率の低下が中心となる失敗、意味・操作・合成の変化に応じた不均一な性能低下が明らかになった。 診断に加えて、GIFTBenchの規模と多様性は、従来のIFLデータセットよりも大幅に広い学習分布を提供する。代表的な位置特定モデルをGIFTBenchで学習すると、外部データセットへの転移を集約した性能が一貫して向上した。このことは、このベンチマークが評価ツールであるだけでなく、領域をまたぐ位置特定のための有効な学習資源にもなることを示している。 さらに診断結果に基づき、検出と位置特定の枠組みForenScopeを開発した。分類に適応させた表現と、複数の深さ・縮尺の空間特徴、学習による層の融合、粗い縮尺の情報を選択的に条件として与える仕組みを組み合わせる。実験では、画像単位の検出能力を維持しながら、データセットをまたぐ位置特定が改善した。GIFTBenchのデータセット紹介ページはhttps://giftbench-preview.doudoudouya337.chatgpt.siteで利用できる。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Reliable evaluation of image forgery localization (IFL) requires assessing models under diverse distribution changes, yet existing benchmarks often cover limited manipulation conditions or entangle multiple factors in cross-dataset evaluation. Consequently, aggregate performance provides an incomplete view of localization generalization. We introduce GIFTBench, a multi-axis benchmark of 115,013 manipulated images with pixel-level annotations spanning manipulation source, semantic target, editing operation, and composition complexity. GIFTBench supports axis-specific transfer analysis and evaluation on twelve external datasets. Its diagnostic studies reveal asymmetric cross-source transfer, recall-dominated failures, and heterogeneous degradation across semantic, operational, and compositional changes. Beyond diagnosis, the scale and diversity of GIFTBench provide a substantially broader training distribution than conventional IFL datasets. Training representative localizers on GIFTBench consistently improves their aggregate transfer to external datasets, showing that the benchmark serves not only as an evaluation tool but also as an effective training resource for cross-domain localization. Guided by the diagnostic findings, we further develop ForenScope, a detection and localization framework combining classification-adapted representations with multi-depth, multi-scale spatial features, learned layer fusion, and selective coarse-scale conditioning. Experiments show improved cross-dataset localization while retaining image-level detection capability. The GIFTBench dataset showcase page is available at https://giftbench-preview.doudoudouya337.chatgpt.site.
arXiv ID: 2610.01778 / 要約の誤りについて