arXiv論文メモ
新着一覧
cs.AI / cs.CC / cs.CE / cs.ET / cs.IR · 査読状況未確認

生成検索の回答で重要な出典関係の欠落を監査

Claim-Gated Source-Risk Auditing for Generative Search

Kainan Zhou, Chuhong Xu, Gangzhen Qian, Zhaoyi Li

この論文をやさしく読む

ひとことで言うと

検索AIの回答が、引用元を正しく示していても重要な関係を省いていないかを調べる監査の仕様です。

何に役立つ?

生成検索の回答を監査する際、証拠不足を『問題なし』と誤認しない判定手順の設計に役立ちます。

この研究の面白いところ

引用の裏付けと出典関係の開示を別々に扱い、81通りの条件と192件の不正記録で仕様どおりの動作を確かめています。

どこまで分かった?

合成テストでの仕様適合を示した段階です。実際の回答に対する検出精度や利用者への効果はまだ評価されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

生成検索の回答は、引用した記述自体に裏付けがあっても、その解釈を変える出典間の関係を省くことがある。本研究では、質問、出典、回答の組を対象に、個々の主張を条件にした監査方法を定める。関係を示す証拠、回答がその主張を採用したこと、関係の重要性、その開示の四つがすべて観察された場合だけ、欠落の判断を確定する。証拠が不完全な場合は、関係がないと扱わず、未解決のままにする。この仕様では、その判断を引用の裏付けの有無やレビューの優先度から分け、版の付いた証拠範囲に判断を結び付ける。参照用の検査器により、記録の形式に関する取り決めを実行可能にする。全組み合わせを網羅した合成テストでは、三状態の判定条件81通りをすべて再現し、意図的に不正な192件の記録を拒否した。共通の条件を持つ比較手法や判定条件の一部を除いた比較によって、最終判断の論理を証拠不足の扱いから切り分け、制御した状態遷移で引用の裏付けとの分離と証拠除去時の挙動を確認した。これらは有限個の条件について仕様への適合を確認した結果であり、検出精度や利用者にとっての結果改善を示すものではない。意味的な妥当性と実運用での利点を確立するために必要な、独立した注釈付け、未使用データでの評価、対応のある有用性試験も定義する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

A generative search answer can cite a supported passage yet omit a source relationship that changes its interpretation. We specify a claim-gated audit of the query-source-answer tuple. An omission is resolved only when relationship evidence, answer adoption, materiality, and disclosure are all observed; incomplete evidence remains unresolved rather than being treated as independence. The specification separates this endpoint from citation support and review priority, and binds decisions to versioned evidence spans. A reference checker makes the record contract executable. On an exhaustive synthetic suite, it reproduces all 81 three-state predicate combinations and rejects 192 deliberately malformed records. Common-guard baselines and predicate ablations isolate endpoint logic from missing-evidence handling, while controlled transitions check support separation and evidence removal. These are finite contract-conformance results, not detector accuracy or evidence of improved user outcomes. We define the independent annotation, held-out evaluation, and paired utility tests still required to establish semantic validity and deployment benefit.

著者のコメント

International Conference on Artificial Intelligence, Automation and Algorithms (AI2A 2026)

arXiv ID: 2609.29145 / 要約の誤りについて