同じ情報源を使うたびに理解を深める言語モデルエージェント
From Knowledge Access to Source Learning: Developing Source-Specific Competence
この論文をやさしく読む
ひとことで言うと
同じ資料を何度も検索するだけでなく、その構造や使い方についての理解を保存・改善するエージェントです。理解の穴や課題での困り方から、読み直す場所を決めます。
何に役立つ?
継続して同じ資料群を使う知識集約型の作業に向けた方法です。複数のベンチマークとLLMで、検索や従来の記憶方式との比較結果を示しています。
この研究の面白いところ
何を学び直すかは経験から決めても、保存する更新内容は元の情報源から作り直します。経験由来の気付きを、出典に戻した理解へ変える構成です。
どこまで分かった?
最良だったのは15設定中13設定で、最大22.6ポイントはHybrid RAGとの比較における最大改善幅です。全設定の平均改善ではありません。要旨には各指標の定義や再読・更新の計算コストはありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
大規模言語モデル(LLM)エージェントは、知識を多く必要とする一連の課題を解くため、継続的に使う外部情報源への依存を強めている。既存手法は情報源の内容へのアクセスや整理を改善し、エージェントの記憶システムは過去のやり取りから再利用できる知識を保持する。しかし、同じ情報源を繰り返し使うことは、その理解を段階的に深める機会というより、依然として主にアクセスの繰り返しとして扱われている。私たちは、継続して参照する信頼できる情報源について、再利用可能な、その情報源に固有の能力を育てる「情報源学習」を研究する。 この能力を、知識がどう構造化され、解釈され、適用されるかを含め、情報源への再利用可能な理解を捉える永続的な情報源モデルで表す。その構築と段階的な改善のため、相補的な2つの学習機構を組み合わせるSourceLearnを提案する。自己主導の情報源学習は、理解が不完全な部分を特定し、適応的に情報源を読み直す。課題に導かれる情報源学習は、後続課題での経験を使って、局所的な表現の欠落や、情報源の知識をどう整理すべきかに関する繰り返し生じる必要性を明らかにする。どちらも、学習信号によって見直す対象を決める一方、永続的な更新内容は信頼できる元の情報源から再構成する。 5つのベンチマークと3つのLLM基盤にわたり、SourceLearnは15設定中13設定で最高性能を達成した。Hybrid RAGに対する改善は最大22.6ポイントであり、静的な情報源表現や経験ベースの記憶のベースラインに対しても、全体として大きく改善した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Large language model (LLM) agents increasingly rely on persistent external sources to solve sequences of knowledge-intensive tasks. Existing methods improve how source content is accessed and organized, while agent-memory systems preserve reusable knowledge from prior interactions, but repeated use of the same source is still largely treated as repeated access rather than an opportunity to progressively improve understanding of that source. We study source learning: developing reusable source-specific competence over a persistent authoritative source. We represent this competence with a persistent source model that captures reusable understanding of the source, including how its knowledge is structured, interpreted, and applied. To construct and progressively refine such models, we propose SourceLearn, which combines two complementary learning mechanisms. Self-Directed Source Learning identifies what remains incompletely understood and adaptively revisits the source, while Task-Guided Source Learning uses downstream experience to reveal local representational gaps and recurring needs in how source knowledge should be organized. In both cases, learning signals determine what should be reconsidered, while persistent updates are reconstructed from the authoritative source. Across five benchmarks and three LLM backends, SourceLearn achieves the best performance in 13 of 15 settings, with gains of up to 22.6 points over Hybrid RAG and substantial overall improvements over static source representations and experience-based memory baselines.
著者のコメント
Website: https://sourcelearn.github.io/ Code: https://github.com/luchengfu6/SourceLearn
arXiv ID: 2610.02150 / 要約の誤りについて