arXiv論文メモ
新着一覧
cs.HC · 査読状況未確認

自動運転の検証に使う人間らしい歩行者モデル

A Human-Like Pedestrian Model for Automated Driving Simulations

Ruofeng Wang, Patrick Ebel, Philipp Wintersberger and Antti Oulasvirta

この論文をやさしく読む

ひとことで言うと

交通状況に応じて判断を変える歩行者を、自動運転のシミュレーター内で再現するモデルです。

何に役立つ?

自動運転車が歩行者と接する場面の開発や評価で、多様な横断行動を試す用途が考えられます。実際の車両での安全性を実証したという結果ではありません。

この研究の面白いところ

危険の感じ方や時間的圧力をモデルの制約に取り込み、横断判断だけでなく、ためらいや回避時の速度調整も再現しています。

どこまで分かった?

結果はシミュレーター内の学習と、未経験の交通環境への方策の移行について述べています。実交通での運用成績は要旨に示されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

自動運転車は多様な交通状況で歩行者と安全かつ効率よく相互作用する必要がある。運転シミュレーターはその能力を学習するための拡張しやすい試験環境となるが、理論に基づく既存の歩行者モデルは、単一車線で横断するかどうかの判断に対象が限られている。データに基づく手法は複雑な状況での歩行者行動を予測できる一方、まれで安全上重要な場面の観測が十分でない。そこで、多車線、交通量の多い状況、危険な運転スタイルを含む現実的で複雑な交通状況において、人間らしい行動を明確に示す方策をシミュレーター内で学習させる方法を提案する。技術的な貢献は、歩行者と車両の相互作用を、知覚、認知、運動に関する理論に基づく制約を備えた部分観測マルコフ決定過程として新たに定義した点にある。交通中の人間の適応的な性質を取り込み、知覚した危険、時間的圧力、状況の複雑さに応じて反応をどう調整するかを模擬する。シミュレーターでドメインランダム化を伴う深層強化学習によって訓練すると、車間の隙間を見て横断する判断、車の譲りを受け入れる判断、ためらい、危険を避けるための速度調整など、これまでに示された人間の横断行動に関する実証的知見を幅広く再現した。学習した方策は未経験の交通環境にも移り、追加学習によって地域の交通慣習にも適応できる。これらの結果は、自動運転システムの開発と評価を支える、シミュレーターで使用可能な歩行者モデルの設計指針を示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Automated vehicles must be able to interact with pedestrians safely and efficiently across diverse traffic situations. Although driving simulators offer a scalable testbed for learning such capabilities, existing theory-inspired pedestrian models are narrow in scope and limited to go/no-go crossing decisions in single-lane settings. While data-driven approaches can predict pedestrian behavior in complex situations, they lack sufficient observations in rare, safety-critical scenarios. Here, we propose an approach to training pedestrian models in simulators so that learned policies generate demonstrably human-like behavior in realistic, complex traffic scenarios, including multiple lanes, heavy traffic, and dangerous driving styles. Our technical contribution is a novel definition of pedestrian-vehicle interaction as a partially observable Markov decision process (POMDP) with theory-grounded perceptual, cognitive, and motor constraints. It accounts for the highly adaptive nature of human behavior in traffic and simulates how people adjust their responses according to perceived danger, time pressure, and the complexity of the situation. When trained via deep reinforcement learning (RL) with domain randomization in a simulator, the model reproduces the broadest range of empirical findings shown so far on human crossing behavior, including gap acceptance, yielding acceptance, hesitation, and evasive speed adjustment. We show that learned policies transfer to unseen traffic environments, and can be further adapted to local traffic norms with finetuning. Together, these results establish a blueprint for simulator-ready pedestrian models that can support the development and evaluation of automated driving systems.

著者のコメント

14 pages, 7 figures, 1 table. Submitted to IEEE Transactions on Intelligent Transportation Systems

arXiv ID: 2609.29175 / 要約の誤りについて