arXiv論文メモ
新着一覧
cs.CL / cs.LG / q-fin.TR · 査読状況未確認

ニュース取引の言語モデルを費用・遅延・流動性込みで評価

Financial Language Models as Applied Artificial Intelligence Systems for News-Based Trading under Market Frictions

Kemal Kirtac

この論文をやさしく読む

ひとことで言うと

金融ニュースを読むAIの評価を文章分類の精度だけで終えず、情報が届く時刻や売買費用、流動性を含む取引判断までつなげる研究です。

何に役立つ?

ニュースを用いる取引モデルの研究で、後から知った情報の混入や費用の無視を避け、比較を再現しやすくする評価手順として役立ちます。

この研究の面白いところ

モデルの公開時期と学習データの対象期間を意識した標本外評価に加え、公開データでも再現できる評価系を用意しています。予測性能と運用負荷の両方を見ます。

どこまで分かった?

デコーダー型の優位は研究での比較結果です。要旨には収益率、取引期間、運用金額などの具体値がなく、将来の利益や任意の市場での優位性を保証する結果ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

金融言語モデルは、構造化されていない企業固有のニュースを構造化された意思決定シグナルへ変換できる。しかし金融AI研究には、こうしたシグナルが金融の意思決定システムでも有用であり続けるかを評価する統合的な実装枠組みが欠けている。計算機科学では、時系列予測、テキスト分類、マルチモーダル株価予測、グラフに基づく市場モデル化、機械学習の運用について優れた方法が開発されてきた。それでも、金融言語モデルの出力を、イベント時点での情報の観測可能性、確率の較正、執行時刻、取引費用、流動性制約、運用規模の上限、運用診断、統計的推論のすべてを通じて検証する、分野固有の手順は提供されていない。 本研究では、時刻付きの金融テキストを、監査可能で再現可能、かつ市場で実行可能な取引判断へ変換する、市場摩擦を考慮した感情分析から取引への枠組みMFASTを導入する。対象はニュースに基づく取引であり、ポートフォリオの判断を評価する前に、企業固有の文章を証券へ結び付ける必要がある。この枠組みはRefinitiv News AnalyticsをCenter for Research in Security Prices(CRSP)の株式データと結び付ける。主要な標本外評価は、モデル公開後のニュースで、かつ基盤モデルについて公表されたデータの最新時点の対象期間外にあるものに限定する。また、公開の金融テキストと価格データによる公開再現用の評価系を加える。 結果は、デコーダーのみの言語モデルが、分類、較正、リターン予測、費用控除後のポートフォリオ性能において、エンコーダー型のベースラインと辞書ベースの感情分析を上回ることを示す。一方、運用診断では、正確さ、遅延、メモリ、処理量、推論費用の間のトレードオフが明らかになる。本稿は、金融言語モデルを信頼できる形で評価するには、言語理解、厳密な時間管理、市場摩擦を考慮した実装、再現可能な検証を組み合わせた、端から端までの工学的な取り組みが必要であることを示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-20(UTC)
最新改訂
2026-09-20 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Financial language models can transform unstructured firm-specific news into structured decision signals, but financial AI research lacks an integrated deployment framework for evaluating whether those signals remain useful in financial decision systems. Computer science research has developed strong methods for time-series forecasting, text classification, multimodal stock prediction, graph-based market modeling, and machine-learning operations, yet these streams do not provide a domain-specific protocol that jointly tests financial language-model outputs under event-time observability, probability calibration, execution timing, transaction costs, liquidity constraints, capacity limits, operational diagnostics, and statistical inference. We introduce MFAST, a Market-Friction-Aware Sentiment-to-Trading framework that converts timestamped financial text into auditable, reproducible, and market-feasible trading decisions. The application is news-based trading, where firm-specific text must be linked to securities before portfolio decisions can be evaluated. The framework links Refinitiv News Analytics to Center for Research in Security Prices (CRSP) equity data, restricts the primary out-of-sample evaluation to post-release news outside disclosed foundation-model data-freshness periods, and adds a public replication arm using open financial text and public price data. Results show that decoder-only language models outperform encoder baselines and dictionary sentiment in classification, calibration, return prediction, and net portfolio performance, while operational diagnostics reveal trade-offs among accuracy, latency, memory, throughput, and inference cost. The paper shows that credible evaluation of financial language models requires an end-to-end engineering approach combining language understanding, temporal discipline, market-friction-aware deployment, and reproducible validation.

著者のコメント

47 pages. Revise and resubmit at Engineering Applications of Artificial Intelligence

arXiv ID: 2609.23703 / 要約の誤りについて