拡散モデルの画像生成で属性の比率と多様性を同時に調整する
Debias Anything: Fairness with Diversity without Supervision in Diffusion Models
この論文をやさしく読む
ひとことで言うと
画像生成で特定の属性の出現比率を調整しながら、似た画像ばかりになるのを防ぐ方法です。
何に役立つ?
生成する画像群の属性構成を指定した比率へ近づけ、品質と多様性も保つ用途が考えられます。センシティブ属性の注釈付き学習データを用いない方法として提案されています。
この研究の面白いところ
属性比率を誘導する方向と、多様性を測る意味的な不一致を、同じ視覚言語表現から作ります。固定した拡散モデルへアダプターを接続する構成です。
どこまで分かった?
要旨の公平性は、主にバッチ内の属性表現の比率を調整する文脈です。公平性全般を解決するとの結論ではありません。属性の注釈データは不要とされますが、方向を指定するテキストプロンプトは使用します。具体的な実験値は要旨にありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
拡散モデルは高品質な画像を生成する一方、学習データに含まれる人口統計的な不均衡を再現し、増幅する。あるセンシティブ属性に関して、学習後に生成過程の偏りを減らすには、通常は分類器による誘導や明示的な追加テキスト条件を使うが、これらは手法の適用性や出力の多様性を下げる。逆に、多様性だけを促す方法では、属性が公平に表現される保証はない。 本論文では、公平性と多様性に同時に取り組み、一般に任意の拡散モデルとセンシティブ属性へ適用できる手法を提案する。このため、アダプターで固定された拡散モデルを事前学習済みの視覚言語埋め込み空間へ接続し、センシティブ属性の注釈なしで公平性と多様性を誘導する。公平性については、一対のテキストプロンプトが属性方向を定め、バッチの構成を特定の比率へ導く。多様性については、この表現から導かれる意味の推定同士の不一致を測るスコアを導入する。 この定式化は、無条件とテキスト条件付きの拡散モデルに対応し、センシティブ属性に関する事前知識やデータを必要としない。実験では、同程度の公平性水準で、品質と多様性のスコアを改善することを確認した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Although diffusion models produce high-quality images, they also reproduce and amplify demographic imbalances in their training data. Debiasing their generation process post-training w.r.t. some sensitive attribute usually relies on classifier guidance or explicit text extra-conditioning, but this reduces methods' applicability and output diversity. Conversely, methods promoting diversity alone do not ensure fair attribute representation. In this paper, we propose a method tackling fairness and diversity jointly that is generally applicable to any diffusion model and any sensitive attribute. To this end, an adapter connects the frozen diffusion model to a pretrained vision-language embedding space, enabling fairness and diversity guidance without sensitive-attribute annotations. For fairness, pairs of text prompts define attribute directions which guide batch composition towards specific proportions. For diversity, we introduce a score measuring disagreement between the semantic estimates derived from this representation. The formulation supports unconditional and text-conditional diffusion models, while requiring no prior knowledge or data of sensitive attribute. Experiments confirm that our method improves quality and diversity scores at comparable fairness levels.
arXiv ID: 2610.01815 / 要約の誤りについて