arXiv論文メモ
新着一覧
cond-mat.stat-mech · 査読状況未確認

まれな揺らぎを確率最適化で推定するLDSTOPの使い方

Stochastic optimisation method for estimating large deviations

Daniël W. H. Cloete and Hugo Touchette

この論文をやさしく読む

ひとことで言うと

通常は起こりにくい大きな揺らぎを、その揺らぎが起こるよう徐々に制御した確率過程から調べる計算方法です。異なる種類のMarkov過程で使う手順を説明しています。

何に役立つ?

非平衡系のまれな事象を数値的に評価する際、目的関数の作り方やニューラルネットワークでの表現を選ぶ参考になります。既存の機械学習ライブラリと組み合わせる実装も対象です。

この研究の面白いところ

多数の候補軌道を同時に保持する発想ではなく、単一軌道のシミュレーションを少しずつ望む駆動過程へ導きます。離散時間、ジャンプ、拡散という異なるモデルを共通の最適化の考え方で扱います。

どこまで分かった?

要旨では簡単な応用による説明を挙げていますが、計算速度や誤差の具体的な比較数値はありません。任意の高次元問題で効率的に収束する保証まで述べているわけではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

Markov過程としてモデル化した非平衡系の揺らぎを統計物理学で特徴付ける大偏差関数を、効率的に計算する確率最適化手法が最近提案された。LDSTOPと呼ばれるこの手法は、制御理論と機械学習の数値技法を組み合わせ、『駆動過程』を反復的に構成する。駆動過程とは、非平衡系のモデルとして使うMarkov過程を制御したもので、その系の指定された揺らぎ、すなわち大偏差を最適な形で実現する。スペクトル近似、重点サンプリング、クローニング、分岐に基づく他の手法と比べ、LDSTOPは単一の軌道をシミュレーションし、それを徐々に駆動過程へ導くことで動作するため、単純で柔軟性があり、規模を拡張しやすい。 本研究では、応用で考えられるMarkov過程の全種類、すなわち離散時間Markov連鎖、連続時間ジャンプ過程、拡散過程にこの方法を適用し、これらの利点を例示する。各種類について、最適化する目的関数を定義し、ニューラルネットワークなどを使って駆動過程を表現するための複数の選択肢を説明し、簡単な応用を通じて実装の詳細を示す。これらを通じて、確率過程のシミュレーションと高次元最適化問題の解法にPyTorch、TensorFlow、JAXなどの既存の機械学習パッケージを組み合わせたときの、この手法の効率性と使いやすさを示すことを目指す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

A stochastic optimisation method was recently proposed to efficiently compute large deviation functions, used in statistical physics to characterise the fluctuations of nonequilibrium systems, modelled as Markov processes. The method, called LDSTOP, combines numerical techniques from control theory and machine learning to iteratively construct the "driven process", a controlled version of the Markov process used as a model of nonequilibrium system that realises a given fluctuation or large deviation of that system in an optimal way. Compared to other methods based on spectral approximations, importance sampling, cloning or splitting, LDSTOP is simple, flexible, and scalable, as it works by simulating single trajectories that are gradually guided towards the driven process. Here, we illustrate these advantages by applying the method on the full range of Markov processes considered in applications, namely, discrete-time Markov chains, continuous-time jump processes, and diffusion processes. For each type, we define the objective function to be optimised, explain different options available for representing the driven process (using, e.g., neural networks), and provide implementation details through simple applications. With these contributions, we aim to showcase the method's efficiency, as well as its ease of use when combined with available machine learning packages, such as PyTorch, TensorFlow and JAX, for simulating stochastic processes and solving high-dimensional optimisation problems.

著者のコメント

30 pages, 9 figures

arXiv ID: 2609.24473 / 要約の誤りについて