arXiv論文メモ
新着一覧
cs.AI / cs.CE / cs.MA / cs.SI · 査読状況未確認

米国の政策と関係者の行動を結ぶ金融シミュレーション用データ

PAWS: Policy-driven Agentic World Simulation

Tiviatis Sim, Jia Hui Woon, Xinming Gao, Chen Gao, Fengbin Zhu, Zheng Huanhuan, Chua Tat Seng, Kenji Kawaguchi

この論文をやさしく読む

ひとことで言うと

米国の政策ニュースと関係者の行動を日付や情報源で結び、金融シミュレーションを過去の事例に照らせるデータセット。

何に役立つ?

政策への反応を再現するエージェントモデルを、根拠となる記事や市場の動きと照合して評価する用途が考えられる。

この研究の面白いところ

36の政策事例でニュースと6万件超の行動を関連付け、正解率が高くてもまれな行動を見逃すという評価上の問題も示した。

どこまで分かった?

データは米国の36政策事例が対象。要旨の事例研究と再生実験は、将来の政策反応を正確に予測できることを示したものではない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

政策介入の影響は、公共の情報発信、組織の意思決定、関係者の反応を通じて広がる。しかし、金融分野の複数エージェントシミュレーションに使うデータセットでは、これらの過程を時系列の揃った過去の根拠資料と結び付けることが少ない。本研究は、検証済みの米国の金融・経済政策36事例、政策に関連するニュース12,727件、情報源に根拠を持つ関係者の行動65,291件を収めた、Policy-driven Agentic World Simulation(PAWS)データセットを提案する。 各行動を根拠となるニュースに結び付け、相互作用の種類、金融行動の分類と細分類、意味上の属性、外部分類体系への条件付き対応を記録する複数層の事象枠で表す。主体を標準化した組織名に対応付け、行動を日次の市場収益率の文脈に揃えて、政策に関わるエージェントのシミュレーションを再生できるようにした。 層化抽出した2,522件の行動について、独立したAIと人間の評価者は、相互作用の種類の最初の判定で89.4%一致し、不一致は後に裁定した。2008年の空売り禁止と2001年の株価呼値の十進法化を扱う事例研究では、ニュースが多い場合と少ない場合の両方で、文書化された政策の時系列と関連する市場の動きを再現した。再生実験は、高い正解率でもまれな関係者の行動を見落とし得ることを示し、行動の時期と較正が主要な課題だと明らかにした。PAWSは、過去の根拠に基づく金融シミュレーションで、エージェントの影響、政策反応の連鎖、行動と結果の対応を検証可能な形で評価する土台を提供する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Policy interventions propagate through public communication, institutional decisions, and stakeholder responses, yet datasets for financial multi-agent simulation rarely connect these processes to temporally aligned historical evidence. We introduce PAWS, a Policy-driven Agentic World Simulation dataset covering 36 verified U.S. financial and economic policy episodes, 12,727 policy-linked news records, and 65,291 source-grounded stakeholder actions. Each action is linked to its supporting news and represented by a multi-layer event frame capturing its interaction mode, financial-action family and subtype, semantic attributes, and conditional mappings to external taxonomies. Entities are resolved to normalized organizations, and actions are aligned with daily market-return context to support policy-agent simulation replay. On 2,522 stratified action samples, independent AI and human reviewers achieved 89.4% initial agreement on interaction mode, with disagreements subsequently adjudicated. Case studies of the 2008 short-selling ban and 2001 decimalization recover documented policy timelines and associated market patterns across both dense and sparse news settings. A replay study further shows that high accuracy can mask failure to detect rare stakeholder actions, identifying action timing and calibration as central challenges. PAWS provides an auditable substrate for evaluating agent influence, policy-response cascades, and action-outcome alignment in historically grounded financial simulations.

arXiv ID: 2609.28547 / 要約の誤りについて