arXiv論文メモ
新着一覧
cs.LG / cs.AI · 査読状況未確認

時系列基盤モデルの予測精度と確率予測の信頼性

Evaluating Accuracy and Probabilistic Reliability of Zero-Shot Time Series Foundation Models

Panagiotis Michael and Moysis Symeonides and Demetris Trihinas

この論文をやさしく読む

ひとことで言うと

六つの時系列基盤モデルを比較し、予測値の当たりやすさと予測確率の信頼性を調べた研究。

何に役立つ?

エネルギー、交通、金融の時系列予測で、点予測と不確実性のどちらを重視するかに応じたモデル選択の参考になる。

この研究の面白いところ

xLSTMは期間をまたぐ確率校正に強みを示した一方、パッチ型Transformerには長期予測で校正上の課題が見られた。

どこまで分かった?

結果は評価した六つのモデルとデータセットに基づく。要旨には各手法の誤差や校正指標の具体的な数値は示されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

時系列基盤モデル(TSFM)は、課題ごとの学習をなくしたゼロショット予測への転換を期待されている。しかし、既存研究では予測精度と確率予測の校正の両立関係が見過ごされがちである。本論文は、エネルギー、交通、金融のデータセットで六つのTSFMを評価するベンチマーク研究を示し、統計的な基準手法および教師あり深層学習モデルと性能を比較する。TSFMは統計的手法や教師ありモデルより良好な性能を示した一方、点予測の精度と確率予測の信頼性の間に根本的なトレードオフが見られた。具体的には、xLSTM構造は予測期間をまたいで頑健な確率校正を示した。対して、パッチに基づくTransformerは競争力のある精度を示すものの、長い予測期間では校正に問題があった。また、Transformerに基づくモデルには、最適なゼロショット予測のための文脈長が飽和する点が見られた。これらの知見は、実際の運用で汎化性能と不確実性の定量化の釣り合いを取るための根拠を提供する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Time Series Foundation Models (TSFMs) promise a paradigm shift toward zero-shot forecasting by eliminating task-specific training. However, existing works often overlook trade-offs between predictive accuracy and probabilistic calibration. This paper presents a benchmark study of six TSFMs evaluated on energy, traffic, and financial datasets. We contrast their performance against statistical baselines and a supervised DL model. The study reveals that while TSFMs outperform statistical methods and supervised models, they are subject to a fundamental trade-off between point accuracy and probabilistic reliability. Specifically, xLSTM architectures provide robust probabilistic calibration across horizons. In contrast, patch-based transformers offer competitive accuracy but face calibration issues at long horizons, while transformer-based models exhibit context saturation points for optimal zero-shot reasoning. These findings offer evidence-based guidance for balancing generalization and uncertainty quantification in real-world deployments.

著者のコメント

Accepted for publication at the 30th European Conference on Advances in Databases and Information Systems (ADBIS 2026)

arXiv ID: 2609.25788 / 要約の誤りについて