arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

交通シミュレーターの調整と交通制御を共通表現で結ぶ

Learning to Move Cities: Deep Meta-Models and Reinforcement Policies for Calibration and Control in Urban Networks

Adewumi Augustine Adepitan, Christopher J. Haruna, Oluwasegun Adegoke, Ayooluwatomiwa Ajiboye, Oluwatobi Oluwasakin

この論文をやさしく読む

ひとことで言うと

交通モデルを現実のデータに合わせる作業と、渋滞を減らす経路・時間調整を、共通の圧縮表現を使ってつなぐ方法です。

何に役立つ?

都市交通のシミュレーション調整と運用方策の学習を一続きに行うための枠組みになります。実際の交通計画への活用は想定される用途で、要旨の削減率はベンチマークネットワークでの評価です。

この研究の面白いところ

圧縮表現を較正だけで終わらせず、強化学習の状態として再利用します。モデルを合わせる段階と制御する段階が、同じ交通ダイナミクスの表現を共有します。

どこまで分かった?

移動時間の51%削減は最大値であり、平均値ではありません。要旨には実都市での導入試験やネットワークごとの改善幅は示されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

都市交通ネットワークには、高忠実度シミュレーターの較正からリアルタイムの運用制御まで、複雑な最適化の課題がある。本論文は、都市交通ダイナミクスについて共通に学習した表現を介して、シミュレーターの較正と強化学習による制御を結ぶ、共有潜在空間の枠組みを提示する。 まず、多層パーセプトロン(MLP)とオートエンコーダーを組み合わせた構成を開発し、シミュレーターの入力である起終点間需要やネットワークパラメーターと、出力である移動時間や混雑パターンを結び付ける低次元多様体を学習する。これにより、較正のための効率的なベイズ最適化が可能になる。この方法は従来の次元削減法よりも高い標本効率を示し、一定の計算予算の下で観測データへのより良い適合を達成する。 次に、経験再生とターゲットネットワークを備えた深層Q学習エージェントを実装し、スケジュールと経路の調整を通じて動的な交通配分を最適化する。ベンチマークネットワークでの実証的な評価では、基準となる運用と比べ、ネットワーク全体の移動時間を最大51%削減する。学習した潜在表現は、ベイズ較正の次元削減だけでなく強化学習の状態表現にも組み込まれ、制御方策が圧縮され較正された交通ダイナミクスに基づいて動作できるようにする。 この共有潜在空間による定式化は、高度交通システムにおいて、シミュレーターの較正から適応的な運用制御へ至る統一的な道筋を提供する。結果は、特に従来の最適化手法が計算上のボトルネックに直面する大規模ネットワークについて、都市交通の計画と管理を変える深層学習の可能性を示している。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Urban transportation networks present complex optimization challenges spanning calibration of high-fidelity simulators and real-time operational control. This paper presents a shared latent-space framework that connects simulator calibration and reinforcement learning control through a common learned representation of urban traffic dynamics. First, we develop a combinatorial MLP-autoencoder architecture that learns low-dimensional manifolds linking simulator inputs (origin-destination demand, network parameters) to outputs (travel times, congestion patterns), enabling efficient Bayesian optimization for calibration. This approach demonstrates superior sample efficiency compared to traditional dimension reduction methods, achieving better fit to observational data within fixed computational budgets. Second, we implement a deep Q-learning agent with experience replay and target networks to optimize dynamic traffic assignment through scheduling and routing adjustments. In empirical evaluations on benchmark networks, our approach reduces system-wide travel times by up to 51% compared to baseline operations. The learned latent representation is not only used to reduce the dimensionality of Bayesian calibration, but is also incorporated into the reinforcement learning state representation, allowing the control policy to operate on compressed and calibrated traffic dynamics. This shared latent-space formulation provides a unified pathway from simulator calibration to adaptive operational control within intelligent transportation systems. Our results highlight the transformative potential of deep learning methods in urban mobility planning and management, particularly for large-scale networks where traditional optimization approaches face computational bottlenecks.

著者のコメント

7 pages, 2 figures. Accepted for publication in the Proceedings of the 2026 IEEE 29th International Conference on Intelligent Transportation Systems (ITSC), Naples, Italy. (c) 2026 IEEE. Personal use of this material is permitted; permission from IEEE must be obtained for all other uses

arXiv ID: 2609.21945 / 要約の誤りについて