RAGの証拠文脈を可視化して科学文献探索を改善するFootprintRAG
FootprintRAG: Visual Analytics for Evidence Context Refinement in RAG-based Scientific Literature Exploration
この論文をやさしく読む
ひとことで言うと
文献RAGが回答を作る前に、どの根拠を集め、残し、捨てたかを利用者が見て直せる分析画面です。
何に役立つ?
検索で見落とした文章や図を回収し、生成された要約を根拠へたどれるようにする文献探索の支援です。
この研究の面白いところ
文献を文章と図の単位に分け、複数の検索方向と反復の履歴、根拠候補の状態を連動表示します。生成前の根拠集合を修正できる対象にしています。
どこまで分かった?
2事例、利用者研究、代表的RAGとのワークフロー比較で評価しています。要旨には参加者数や定量的な効果量はなく、全回答の正しさを保証するものではありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
検索拡張生成(RAG)は、大規模言語モデル(LLM)の出力を科学文献に根拠付けるために広く使われるようになっている。しかし、開放型の文献探索では、生成に使われる証拠文脈が、隠れた検索、再ランキング、評価、フィルタリングの段階を経て作られることが多い。利用者は検索結果の要約を受け取っても、システムが証拠文脈をどのように構成したのか、どの証拠単位を残しどれを捨てたのか、統合前に役立つ可能性のある証拠が除外されたかを知らない場合がある。 本研究では、RAGによる科学文献探索で証拠文脈を改良するための、LLMエージェントを用いた可視分析システムFootprintRAGを提案する。中心となる考えは、RAGの証拠文脈を、生成前に明示的に検査・修正できる分析対象として扱うことである。FootprintRAGは、科学文献を文章と図の証拠単位に分解し、初期クエリを複数の並列なクエリ候補へ展開し、反復ラウンドを通じて証拠を検索・評価する。また、コーパス全体の証拠空間からERSで順位付けした補助候補を提示する。連携したビューによって、検索の軌跡、証拠状態の改訂、出典を追跡できる要約生成を、利用者が操作できるワークフローとしてつなぐ。 ケーススタディ2件、ユーザー研究、代表的なRAGシステムとのワークフローレベルの比較でFootprintRAGを評価した。結果は、検索方向の比較、候補証拠の修正、見落とされた可能性のある証拠の回収、生成要約からそれを支える証拠単位への追跡をFootprintRAGが助けることを示した。FootprintRAGは https://github.com/meteorshowering/FootprintRAGVA.git で公開されている。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-17(UTC)
- 最新改訂
- 2026-09-17 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-17 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Retrieval-Augmented Generation (RAG) is increasingly used to ground large language model (LLM) outputs in scientific literature. However, in open-ended literature exploration, the evidence context used for generation is often produced through hidden retrieval, reranking, assessment, and filtering steps. Users may receive retrieval summaries without knowing how the system constructed the evidence context, which evidence units were retained or discarded, or whether potentially useful evidence was excluded before synthesis. We present FootprintRAG, an LLM-agent-powered visual analytics system for evidence context refinement in RAG-based scientific literature exploration. The core idea is to treat the RAG evidence context as an explicit, inspectable, and revisable analytical object before generation. FootprintRAG parses scientific literature into text and figure evidence units, expands an initial query into parallel query variants, retrieves and assesses evidence across iterative rounds, and surfaces ERS-ranked supplementary candidates from the corpus-level evidence space. Through coordinated views, the system connects retrieval trajectories, evidence-state revision, and provenance-aware summary generation into a user-steerable workflow. We evaluate FootprintRAG through two case studies, a user study, and a workflow-level comparison with representative RAG systems. The results show that FootprintRAG helps users compare retrieval directions, revise candidate evidence, recover potentially overlooked evidence, and trace generated summaries back to supporting evidence units. FootprintRAG is available at https://github.com/meteorshowering/FootprintRAGVA.git.
arXiv ID: 2609.19601 / 要約の誤りについて