歴史資料の出典を追跡できる検索エージェントTRACE
TRACE: Accountable Agentic Retrieval for Source Discovery in Digital Archives
この論文をやさしく読む
ひとことで言うと
OCRの誤りや文書形式の違いがある歴史資料から、出典を追える形で関連文書を探す検索エージェントです。
何に役立つ?
議会記録と新聞などを横断する史料調査を支援します。研究者が根拠となる資料を確認できることを重視し、プロジェクト内では六機関の24人が利用可能な状態です。
この研究の面白いところ
1887年のフランスの議会・新聞を対象とする1,752問で、R@10は0.856、MRRは0.653でした。複数の資料をつなぐ質問で改善が大きく、専用の追加学習なしで取り組んでいます。
どこまで分かった?
評価は特定の歴史資料群に基づきます。1問約0.02ドルという費用も既定のホスト型推論構成での値であり、他の資料や構成で同じ精度・費用になるとは限りません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
歴史資料のアーカイブは、検索拡張生成システムにとって難しい検索対象である。文書はOCRによって品質が損なわれ、ジャンルや出典による異質性が大きく、学術利用や機関での利用には厳密な出典追跡が求められる。歴史コーパスから説明責任を果たせる出典発見を行うため、追加学習を必要としないエージェント型検索の枠組みTRACEを導入する。 本システムは、フランス第三共和政期に議会討論と新聞の間で政治言説が流通する過程を扱う学際プロジェクトDECIDONの中で開発した。同プロジェクトは、デジタル化された歴史資料群と機関での利用事例を対象とする。試作システムは現在プロジェクト内部で稼働し、6つの連携機関に属する24人の研究者が利用できる。 フランス国立図書館のデジタル化資料に由来する1887年の議会討論と新聞を対象に、フランスの歴史に関する1,752問を収めたHistoriQA-ThirdRepublicベンチマークで評価した。TRACEはR@10=0.856、MRR=0.653を達成し、疎検索、密検索、グラフベース、エージェント型RAGの比較手法を上回った。改善は複数段階の推論を要する質問と、複数コーパスにまたがる質問で最も大きかった。 標準のホスト型推論設定では1問当たり約0.02米ドルであり、高価なローカルGPU基盤に依存できない文化遺産機関、研究室、企業にとっても経済的に実行可能である。これらの結果は、大規模なデジタル図書館やアーカイブにおいて、検索の説明責任とコーパスの特性を考慮したエージェント設計が、より重い学習やグラフ構築に基づく方法の実用的な代替となり得ることを示唆する。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-17(UTC)
- 最新改訂
- 2026-09-17 · v1
- 査読・掲載
- 掲載先の記載あり
著者による掲載先の記載:35th ACM International Conference on Information and Knowledge Management (CIKM 2026), Nov 2026, Rome, Italy。出版社での独立確認は未実施です。
更新履歴
- v1 2026-09-17 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Historical archives pose a difficult retrieval problem for retrievalaugmented generation systems: documents are OCR-degraded, heterogeneous across genres and sources, and require strong source traceability for scholarly and institutional use. We introduce TRACE, a training-free agentic retrieval framework designed for accountable source discovery over historical corpora. The system was developed in the context of DECIDON, an interdisciplinary project on the circulation of political discourse between parliamentary debates and the press during the French Third Republic, involving digitised historical collections and institutional use cases. The prototype is currently deployed internally within the project and accessible to 24 researchers across six partner institutions. We evaluate TRACE on HistoriQA-ThirdRepublic, a benchmark of 1,752 French historical questions over parliamentary debates and newspapers from 1887, with documents derived from Bibliothèque nationale de France digitised collections. TRACE achieves R@10 = 0.856 and MRR = 0.653, outperforming sparse, dense, graph-based, and agentic RAG baselines, with the largest gains on multi-hop and cross-corpus questions. At approximately $0.02 per question under the default hosted inference configuration, TRACE also remains economically feasible for heritage institutions, laboratories or companies that cannot rely on costly local GPU infrastructure. These results suggest that, for large digital libraries and archives, retrieval accountability and corpus-aware agent design can provide a practical alternative to heavier training-based or graph-construction approaches.
arXiv ID: 2609.19897 / 要約の誤りについて