arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

周波数特性を保って時系列を生成する潜在フローモデル

Time series generation with spectrally aligned latent flow matching

Camilo Carvajal Reyes and Felipe Tobar

この論文をやさしく読む

ひとことで言うと

時系列を圧縮してから生成すると失われがちな周期や滑らかさを、信号変換に基づく学習で保とうとする研究です。

何に役立つ?

学習用の合成時系列を作る際に、元データに必要な性質が残るかを調べる方法として役立ちます。

この研究の面白いところ

時刻ごとの値の近さだけでなく、周波数や局所構造を表す複数の変換を損失に用います。何を揃えているかを解釈しやすい点が特徴です。

どこまで分かった?

要旨は実データの単変量・多変量ベンチマークでの優位性を述べますが、改善幅やデータセット名は示していません。あらゆる下流の学習課題で性能が改善するという実証でもありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

潜在フローモデルは、時系列生成の信頼性が高く費用効率のよい方法であることが示されてきた。しかし、潜在表現への圧縮は、元のデータセットとのスペクトルの不一致などの望ましくない人工的な特徴を生み、学習用の代替データとしての利用を妨げる。本論文では、スペクトルを整合させた潜在フロー型の時系列生成器を提案する。フローマッチングに用いる潜在空間を、合成標本の適切さに関わる動的特性を保つよう学習させる。 フーリエ変換、ウェーブレット変換、シグネチャ変換などの標準的な信号表現に基づく微調整の損失を取り入れることで、これらの問題を軽減できることが分かった。これらの変換は解釈可能であるため、点ごとの再構成損失だけに頼るのではなく、滑らかさや対象とするスペクトル成分など、重要な特徴について合成信号を実信号に整合させられる。 提案する整合化モデルを、実世界の長い時間範囲の単変量・多変量ベンチマークデータセットで、基本の潜在フローモデルおよび最先端手法と比較する。定量的結果は、学習集合と局所構造を整合させながら、信号の実らしさを反映する指標と計算効率の点で提案法が優れていることを裏づける。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Latent flow models have proven to be a reliable and cost-effective method for time series generation. However, the latent compression induces unwanted artefacts, such as a spectral mismatch with respect to the underlying dataset, thus hindering their use as training surrogates. In this article, we propose a spectrally-aligned latent-flow time series generator, where the latent space for flow matching is trained to preserve dynamical properties that are relevant for the suitability of synthetic samples. We find that incorporating fine-tuning losses based on canonical signal representations such as the Fourier, wavelet and signature transforms helps overcome these issues. The interpretability of these transformations allows us to ensure that the synthetic signals are aligned with the true ones in terms of relevant features, such as smoothness or targeted spectral content, as opposed to relying on pointwise reconstruction losses only. We compare the proposed aligned models against a base latent-flow model and the state of the art over real-world long-range univariate and multivariate benchmark datasets. Our quantitative results validate the superiority of the proposed method in terms of its performance on metrics reflecting signal realness and computational efficiency, while being aligned to the training set with respect to its local structure.

arXiv ID: 2609.21989 / 要約の誤りについて