古いデータを忘れても事前知識を保つ予測制御
Forgetting While Remembering, an Invariant Online Data-Driven Predictive Control Formulation
この論文をやさしく読む
ひとことで言うと
変化するシステムを制御するために古い観測の重みを下げても、安定性などの事前知識は薄れないようにする。
何に役立つ?
ノイズが大きく、対象の性質も時間とともに変わる場合の予測制御を設計する助けとなる。
この研究の面白いところ
変化への追従速度をオンライン調整しつつ、事前分布を不変にする状態方程式で「忘れる対象」を分けている。
どこまで分かった?
検証対象は時変の2次システムである。要旨には追従誤差の数値、比較手法の詳細、物理装置かシミュレーションかの区別は示されていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
線形時変システムに対するオンラインのデータ駆動型予測制御(DPC)では、信号対雑音比(SNR)の低いデータが中心的な課題となる。本論文は、外生入力付き自己回帰モデル(ARX)に基づくベイズ的なオンラインDPCの枠組みを提案する。外部から与えられる事前分布は、システム動特性の滑らかさや安定性などの帰納バイアスを符号化し、SNRが低いときの性能を守るために用いられる。 ARXパラメータの事後推定は、適応速度のハイパーパラメータによって定義された状態方程式を持つカルマンフィルタで時間方向に伝播する。このハイパーパラメータをオンラインで調整し、対象システムの動特性が変化する速さを追跡する。カルマンフィルタの事後平均と共分散は、二次コスト関数の事後期待値である、DPCの最終制御誤差コスト関数を決める。重要なのは、ハイパーパラメータの適応にかかわらず、事前分布が時間を通じて不変に保たれるよう、カルマンフィルタの過程方程式を選ぶ点である。そのため古いデータは忘れても、事前分布は忘れない。時変の2次システムに対するDPC追従実験によって、提案手法の有効性を示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-16(UTC)
- 最新改訂
- 2026-09-16 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-16 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Low signal-to-noise ratio (SNR) data is a core challenge of online Data-Driven Predictive Control (DPC) for linear, time-varying systems. This paper proposes a Bayesian, online DPC framework based on autoregressive models with exogenous inputs (ARX) that uses an externally-provided prior, which encodes inductive bias such as smooth system dynamics and stability, to safeguard performance when SNR is low. The posterior estimate of the ARX parameter is propagated forward in time using a Kalman filter with a state equation defined by an adaptation-rate hyperparameter, which is adjusted online to track the rate at which the underlying system dynamics evolve. The Kalman filter's posterior mean and covariance determine the DPC's Final Control Error cost function, which is the posterior expectation of the quadratic cost function. Critically, the Kalman filter's process equation is chosen so that, regardless of the hyperparameter adaptation, the prior distribution is invariant over time. Thus, while old data is forgotten, the prior is not. DPC tracking experiments on a time-varying second-order system demonstrate the efficacy of the proposed method.
著者のコメント
8 pages, 3 figures, conference paper (CDC)
arXiv ID: 2609.18827 / 要約の誤りについて