連続状態と単語列を往復させる拡散言語モデル
Hierarchical Continuous Diffusion Language Models
この論文をやさしく読む
ひとことで言うと
文章や解答を表す連続的な内部状態を更新しながら、毎回トークン列を読み出して次の更新に戻すモデルです。並列生成でもトークン同士の関係を保つことを狙います。
何に役立つ?
複数の位置を同時に決めつつ、全体の制約を満たす生成方法を研究するのに役立ちます。数独、数の組み合わせの計画、言語モデリングで改善が報告されています。
この研究の面白いところ
連続状態と離散トークンを別々の生成過程として持つのではなく、連続状態だけを保持し、トークンを各段階の手掛かりとして戻します。学習目的もトークン尤度の変分境界から導かれています。
どこまで分かった?
比較はモデル規模を合わせた拡散モデル間で行われています。要旨には改善量、計算時間、自己回帰モデルとの性能比較は記載されていません。言語モデリングでの改善指標は生成パープレキシティです。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
離散拡散言語モデルは、双方向の推論や全体的な制約の充足を必要とする課題で、自己回帰生成の有力な代替となる。しかし、構造上の共通のボトルネックがある。並列に復号するとき、各トークンはそれぞれの周辺分布から独立に抽出されるため、同時に復号するトークン間の統計的な依存関係が断ち切られる。連続拡散言語モデルは、共有する連続状態のノイズを除去することでこの問題を回避するが、ノイズ除去器が見るのはその状態だけであり、最後に復号されるまで、有効なトークン配置と結び付けるものがない。 この問題に対処するため、階層型連続拡散言語モデルHierarchical Continuous Diffusion Language Models(HC-DLM)を提案する。HC-DLMは、離散トークン生成と連続潜在軌道を、原理に基づく単一のノイズ除去過程の中で結び付け、その学習目的関数をトークン尤度の変分境界から導く。独立して完結する離散連鎖に連続的な文脈を付け加える最近の手法とは異なり、HC-DLMでは潜在表現だけを継続的に保持される生成状態とする。各ステップでそこからトークンを読み出し、次の潜在状態更新の足場としてフィードバックする。 構造化推論のSudoku、数学的計画のCountdown、言語モデリングのLM1Bにおいて、同じモデル規模の離散拡散・連続拡散のベースラインを上回った。改善した指標は、SudokuとCountdownではパズルの正答率、LM1Bでは生成パープレキシティである。プロジェクトページ:https://hc-dlm.github.io/ 。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Discrete diffusion language models offer a compelling alternative to autoregressive generation for tasks demanding bidirectional reasoning and global constraint satisfaction. Yet they share a structural bottleneck: when decoding in parallel, each token is sampled independently from its marginal, severing the statistical dependencies among the tokens decoded together. Continuous diffusion language models avoid this by denoising a shared continuous state, but their denoiser sees only that state, so nothing ties it to a valid token configuration until it is finally decoded. To address this, we propose Hierarchical Continuous Diffusion Language Models (HC-DLM), which couple discrete token generation with a continuous latent trajectory in a single, principled denoising process, whose training objective is derived from a variational bound on the token likelihood. In contrast to recent methods that attach continuous context to a self-contained discrete chain, HC-DLM makes the latent the only persistent generative state: tokens are read out from it at every step and feed back as a scaffold for the next latent update. On structured reasoning (Sudoku), mathematical planning (Countdown) and language modeling (LM1B), HC-DLM improves over discrete and continuous diffusion baselines at matched model size, in puzzle accuracy on Sudoku and Countdown and in generative perplexity on LM1B. Project page: https://hc-dlm.github.io/.
arXiv ID: 2610.02193 / 要約の誤りについて