arXiv論文メモ
新着一覧
cs.CV · 査読状況未確認

生成手法ごとの痕跡の違いを残して偽画像を検出する

Revisiting Cross-Reconstruction for Generalizable Deepfake Detection

Bingjian Yang, Shilei Zhao, Zheng Wang

この論文をやさしく読む

ひとことで言うと

偽画像を作った方法ごとの痕跡の違いを消さずに学び、未知の生成方法にも対応しやすい検出器を目指す研究です。

何に役立つ?

学習時と異なる生成器で作られた画像を調べる鑑識の用途に役立つ可能性があります。複数のデータセットと生成器をまたぐ評価で改善を報告しています。

この研究の面白いところ

生成器ごとの違いを邪魔なばらつきと見なすのではなく、複数の手掛かりとして利用しています。意味内容から痕跡を分けつつ、再構成には痕跡も使います。

どこまで分かった?

改善率や評価データセットの名称は要旨にはありません。未知のすべての偽造や実運用で確実に検出できると示したわけではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

既存の画像偽造検出器は、別の条件でも使える鑑識上の手掛かりを捉える能力が限られるため、未知の改変方法への汎化に問題を抱えることが多い。近年の交差再構成に基づく手法は、意味情報と生成痕跡を分離して汎化を改善しようとする。しかし通常、生成器ごとに異なる痕跡を整合させ、再構成の際には痕跡表現を除外するため、改変痕跡に本来ある多様性や視覚的手掛かりを見落とす可能性がある。 本研究では交差再構成を再検討し、頑健な画像偽造検出のための、痕跡を重視した分離の枠組みを導入する。異なる生成過程が生む改変痕跡の内在的な変動、すなわち「痕跡の多様性」は、望ましくない領域差ではなく、互いを補う鑑識上の手掛かりを含むと考える。痕跡を明示的に整合させる代わりに、意味を整合させた生成器間の交差再構成を通じ、多様な痕跡の特徴を保つ。さらに、再構成過程に痕跡表現を取り入れ、意味情報の干渉を減らしながら改変に関連する残差を強調する、マスクを使った周波数考慮型の再構成戦略を導入する。 この設計により、多様な痕跡から、別の条件にも移せる鑑識表現を学べる。複数の基準データセットにわたる広範な実験で、データセットをまたぐ評価と生成器をまたぐ評価の両方で改善を示した。追加分析と要素を取り除く比較実験により、痕跡の多様性の保持と、痕跡を考慮した交差再構成の有効性を検証した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-10-01(UTC)
最新改訂
2026-10-01 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Existing image forgery detectors often suffer from generalization to unseen manipulation methods due to the limited ability to capture transferable forensic cues. Recent cross-reconstruction based methods attempt to improve generalization through semantic-artifact disentanglement, but typically align heterogeneous artifacts across generators and exclude artifact representations during reconstruction, which may overlook the inherent diversity and visual cues of manipulation artifacts. In this work, we revisit cross-reconstruction and introduce an artifact-oriented disentanglement framework for robust image forgery detection. We argue that \textbf{artifact diversity}, i.e., the intrinsic variations of manipulation artifacts introduced by different generation processes, contains complementary forensic cues rather than undesirable domain variations. Instead of enforcing explicit artifact alignment, our framework preserves diverse artifact characteristics through semantically aligned cross-generator reconstruction. Furthermore, we incorporate artifact representations into the reconstruction process and introduce a masked frequency-aware reconstruction strategy to emphasize manipulation-related residuals while reducing semantic interference. This design enables the model to learn transferable forensic representations from diverse artifacts. Extensive experiments on multiple benchmark datasets demonstrate improvements under both cross-dataset and cross-generator evaluation settings. Further analysis and ablation studies validate the effectiveness of artifact diversity preservation and artifact-aware cross-reconstruction.

arXiv ID: 2610.01544 / 要約の誤りについて