指示理解と外部乱数で言語モデルの生成を多様化する
Gacha Decoding: Eliciting Diverse Generations Through Instruction Following
この論文をやさしく読む
ひとことで言うと
出力の単語選択を揺らすだけでなく、外部乱数を使って異なる回答方針を計画させる手法です。
何に役立つ?
考えられる用途は、一定の品質を保ちながら、創作案や設計候補を幅広く集めることです。要旨では複数の生成領域で評価しています。
この研究の面白いところ
高性能化すると回答が似通うという傾向に対し、指示をよく守る能力そのものを多様性の源にする点が特徴です。
どこまで分かった?
最大2.4倍、11.0倍は報告された評価条件での数値です。要旨には各領域の品質判定方法や、タンパク質候補の実験的検証についての説明はありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
モデルの能力向上に応じて効果が伸びる、言語モデルの多様な生成を引き出す推論時手法Gacha Decodingを提案する。実際のチャット、創作、画像生成の計画、タンパク質設計という自由度の高い領域で、同じ品質なら既存の生成多様化手法を大幅に上回る。Vendiは従来手法の次点に対して最大2.4倍となり、同数の高品質な生成モードに、1桁以上少ないサンプル数、具体的には11.0倍のサンプル効率で到達する。また、他のどの手法も見いださなかった新しいモードを発見する。 中心的な着想は、多様性を指示追従の問題として扱うことである。言語モデルのトークンエントロピーに頼るのではなく、その指示追従能力と外部乱数生成ツールのランダム性を組み合わせ、応答空間の異なるモードを拡張性のある形で特定し、実現する。この「サイコロを使って計画する」方法によって、長く観察されてきた多様性とモデル能力の緊張関係を逆転できる。基盤となる言語モデルの指示追従能力が高まると、トークンエントロピーや従来手法での多様性が低下する場合でも、Gacha Decodingによる多様性は一貫して向上する。これらの結果は、トークンエントロピーだけでなく、指示追従によっても生成の多様性を高められることを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We introduce Gacha Decoding, an inference-time method for eliciting diverse language model generations that scales with model capability. Across open-ended domains (in-the-wild chat, creative writing, planning for image generation, and protein design), Gacha Decoding significantly outperforms existing generation diversity approaches at equal quality (up to 2.4x Vendi over the next-best prior approach), reaching the same number of high-quality modes with over an order of magnitude fewer samples (11.0x) and discovering novel modes that no other approach surfaces. Our key insight is to treat diversity as an instruction-following problem: rather than relying on the LM's token entropy, we combine its instruction-following capability with randomness from an external RNG tool to scalably identify and realize distinct modes of the response space. This approach of "planning with dice" enables Gacha to invert the long-observed tension between diversity and model capability. As the underlying LM becomes a better instruction follower, diversity under Gacha Decoding consistently improves--even as its token entropy and diversity under prior approaches decline. Together, our results highlight that instruction following, rather than token entropy alone, can drive generation diversity.
arXiv ID: 2610.01382 / 要約の誤りについて