arXiv論文メモ
新着一覧
cs.CV · 査読状況未確認

1つの固定サイズ網で画像生成の概念学習と位置指定を続ける

CLASP: Continual Low-rank Adapters for Spatially Placed Concepts from One Hypernetwork

Wojciech Gromski, Patryk Krukowski, Jan Miksa, Maciej Zieba, Przemysław Spurek

この論文をやさしく読む

ひとことで言うと

画像生成で新しい概念を順に覚えさせる際、1つの固定サイズのネットワークから概念別の調整を作り、出す位置も指定できるようにします。

何に役立つ?

考えられる用途は、多くの概念を継続して追加する画像生成の個別化です。実験では既存概念の保持と指定位置への配置を評価しています。

この研究の面白いところ

概念ごとに大きな追加モデルや位置制御部品を持つ代わりに、共通のハイパーネットワークから必要な適応を生成します。過去データのリハーサルも不要としています。

どこまで分かった?

コンパクトな概念表現は別途必要なので、総保存量が完全に一定という主張ではありません。実験の概念数、評価指標、比較性能の具体的な数値は要旨にはありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

テキストから画像を生成する拡散モデルを継続的に個別化するには、以前に学んだ概念を保持しながら、新しい概念を順次獲得する必要がある。しかし既存手法は、破滅的忘却が起こるか、概念ごとの追加パラメータや位置に関する構成要素の保存に頼るため、概念が次々に加わるにつれてパラメータの保存量が増える。その結果、個別化タスクの長い系列へ拡張する能力が制限される。 本研究は、サイズが固定された単一のハイパーネットワークを使い、凍結した拡散モデルを継続的に個別化する、過去データのリハーサルを必要としない手法を提案する。新たな概念を獲得するたびにモデルを拡張する代わりに、ハイパーネットワークが、以前に学んだ概念を保持しつつ、個別化に必要な概念固有の適応を動的に生成する。さらにこの枠組みは位置制御を個別化の過程へ組み込み、概念ごとの構成要素を追加せずに、個別化した概念を画像内のどこへ出すかを利用者が指定できるようにする。 この定式化により、コンパクトな概念表現を除けば、学習した概念数に依存しないパラメータ保存量で継続的な個別化が可能になる。実験では、以前に学んだ概念をよく保持し、指定位置との対応付けも安定しており、個別化タスクの長い系列へ効果的に拡張しながら、既存手法と同等以上の性能を示した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-10-01(UTC)
最新改訂
2026-10-01 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Continual personalization of text-to-image diffusion models requires sequentially acquiring new concepts while retaining previously learned ones. However, existing methods either suffer from catastrophic forgetting or rely on storing additional concept-specific parameters and spatial components, causing their parameter footprint to grow with the concept stream. This limits their ability to scale to long sequences of personalization tasks. We propose a rehearsal-free approach that uses a single fixed-size hypernetwork to continually personalize a frozen diffusion model. Instead of expanding the model as new concepts are acquired, the hypernetwork dynamically produces the concept-specific adaptations required for personalization while preserving previously learned concepts. Our framework further integrates spatial control into the personalization process, allowing users to specify where a personalized concept should appear without introducing additional per-concept components. This formulation enables continual personalization with a parameter footprint that remains independent of the number of learned concepts, aside from compact concept representations. Experiments demonstrate strong retention of previously learned concepts and reliable spatial grounding, matching or improving upon existing methods while scaling effectively to long streams of personalization tasks.

著者のコメント

31 pages. Code: https://github.com/genwro-ai/clasp, project page: https://genwro-ai.github.io/clasp

arXiv ID: 2610.01331 / 要約の誤りについて