対数凹でない分布も扱うランジュバン法の収束解析
A Unified Framework for Wasserstein Convergence of ULMC Methods beyond Log-Concavity: Old and New
この論文をやさしく読む
ひとことで言うと
高次元分布から標本を得るLangevin法を共通の予測・修正形式にまとめ、低コストの新方式を導きます。
何に役立つ?
統計計算や機械学習で、標本生成に必要な勾配評価や乱数生成を減らす用途が考えられます。
この研究の面白いところ
一反復で勾配評価1回とGaussian乱数2つを使う方式を導き、従来方式も含めて長時間の誤差を共通に解析します。
どこまで分かった?
収束率は滑らかさなどの条件付きです。非対数凹の場合のW1評価と、強凸の場合のW2評価を同一の前提の保証として読むべきではありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
高次元の確率分布からのサンプリングは、計算統計、科学計算、機械学習に共通する基本課題として、近年ますます注目されている。数多くの手法の中で、減衰不足ランジュバン動力学(ULD)に基づく減衰不足ランジュバン・モンテカルロ(ULMC)法は、効率的な一群として現れている。 本研究では、手法のパラメーターの選び方によって、オイラー型、UBU型、ランダム化スキームをつなぐ「普遍的」予測子・修正子の定式化を導入する。この積分法から、低コストのランダム化積分法(LC-RI)と低コストUBU積分法(LC-UBUI)という二つの新しいクラス、さらに多項式近似と有理近似に基づく、指数関数を使わない変種が得られる。新たなUBU型とランダム化スキームは、1反復につき勾配評価1回とガウス乱数2個だけを必要とし、既存の対応手法に比べ、勾配評価またはガウス乱数の必要数を大きく減らす。 さらに、確率距離における一般的な離散化スキームの長時間誤差解析の枠組みを構築する。一定の滑らかさと非対数凹性の条件の下、この統一的な枠組みを使って、新旧のスキームの非漸近的W₁誤差上限を確立する。収束率は、オイラー型ではO(d^(1/2)h)、UBU型ではO(dh²)、ランダム化型ではO(d^(1/2)h^(3/2))となる。強凸の設定では、同じ非漸近的な誤差上限をW₂距離でも回復できる。数値実験は理論的知見を裏付ける。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-17(UTC)
- 最新改訂
- 2026-09-17 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-17 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
As a fundamental task across computational statistics, scientific computing and machine learning, sampling from high-dimensional probability distributions has received increasing attention in recent years. Numerous sampling algorithms have been proposed, among which underdamped Langevin Monte Carlo (ULMC) methods based on underdamped Langevin dynamics (ULD) have emerged as a class of efficient ones. In this work, we introduce a ``universal" predictor-corrector formulation that bridges Euler-type, UBU-type and randomized schemes through different choices of method parameters. Notably, the ``universal" integrator induces two novel classes of low-cost integrators, termed low-cost randomized integrators (LC-RIs) and low-cost UBU integrators (LC-UBUIs), as well as their exponential-free variants based on polynomial and rational approximations. The resulting new UBU-type and randomized schemes require only one gradient evaluation and two Gaussians per iteration, considerably reducing the number of gradient evaluations or Gaussians per iteration required by existing counterparts. Further, a general framework of long-time error analysis is developed for general discretization schemes in a probability metric. Under certain smoothness and non-log-concavity conditions, we rely on the unified framework to establish non-asymptotic $\mathcal{W}_1$-error bounds of both old and new schemes, revealing convergence rates of order $\mathcal{O}(d^{\frac{1}{2}}h)$ for Euler-type schemes, order $\mathcal{O}(d h^2)$ for UBU-type ones and order $\mathcal{O}(d^{\frac{1}{2}}h^{\frac{3}{2}})$ for randomized ones. In the strongly convex setting, the same non-asymptotic error bounds can be recovered in $\mathcal{W}_2$-distance. Numerical experiments corroborate the theoretical findings.
arXiv ID: 2609.20713 / 要約の誤りについて