arXiv論文メモ
新着一覧
cs.AI · 査読状況未確認

検索拡張生成の研究を効率・防御・対話・推論で整理

Mapping the RAG Landscape: A Four Axis Taxonomy of Efficiency, Defense, Interactivity, and Reasoning

Meghana Sunil, Shravya V, Shravan Venkatraman, Joe Dhanith PR

この論文をやさしく読む

ひとことで言うと

外部の文書を検索して回答するRAGの研究を、速さ、安全性、利用者とのやり取り、複雑な推論の四つの観点で整理したレビューです。

何に役立つ?

RAGシステムを設計するときに、検索精度だけでなく応答速度や防御、対話、推論のどこを改善するか検討するための見取り図になります。

この研究の面白いところ

基本的な構成の分類に加えて、運用で求められる能力を軸に研究を横断整理しています。検索法、融合、埋め込み、強化学習による検索方策も同じ枠組みで扱います。

どこまで分かった?

サーベイであり、新しい一つのシステムの性能を実験で示す要旨ではありません。検索品質や信頼性などは未解決の課題として挙げられており、RAGで誤情報がなくなると結論しているわけではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

大規模言語モデル(LLM)は多くの課題で優れた流暢さを示してきたが、パラメーターに閉じ込められた静的な知識と、誤った情報を生成しやすい性質による制約が残る。検索拡張生成(RAG)は、生成過程へ外部検索を組み込み、モデルの出力を検証可能で最新の情報源に基づかせることで、これらの問題に対処する。従来のサーベイが主にRAGの基本的な構造と標準的な処理過程に注目してきたのに対し、近年の研究は、こうした基礎設計を超えた広い課題と能力を探究している。 本サーベイでは、現代のRAGの発展を統合的かつ構造的に検討し、検索効率の改善、頑健性とセキュリティーの強化、ユーザー主導の対話的な作業過程への対応、多段階または複雑な推論の実現、という四つの軸で分野を分類する。RAGの主要な構成要素を形式化し、密な検索と疎な検索、融合戦略、埋め込みの最適化、強化学習に基づく検索方策にわたる手法を概観し、これらの進展が実際の導入やシステム設計へどう影響するかを明らかにする。評価の実践、分野固有の応用、Naive、Advanced、Modular RAGなどの構造上の派生形も整理する。最後に、検索品質、信頼性、領域適応、拡張性、説明可能性に関して残る課題を概説し、より信頼でき、適応性があり、透明なRAGシステムを構築する機会を示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-10-01(UTC)
最新改訂
2026-10-01 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Large Language Models (LLMs) have demonstrated remarkable fluency across many tasks but remain limited by their static, parameter bound knowledge and their susceptibility to hallucinating information. Retrieval Augmented Generation (RAG) addresses these issues by incorporating external retrieval into the generation process, grounding model outputs in verifiable and up to date sources. While prior surveys primarily focus on core RAG architectures and standard pipelines, recent research explores broader challenges and capabilities that extend beyond these foundational designs. This survey provides a consolidated and structured examination of contemporary RAG developments, organizing the field into a four axis taxonomy: improving retrieval efficiency, strengthening robustness and security, supporting user driven and interactive workflows, and enabling multi step or complex reasoning. We formalize key components of the RAG framework and review methods spanning dense and sparse retrieval, fusion strategies, embedding optimizations, and reinforcement learning based retrieval policies, highlighting how these advances influence practical deployment and system design. We also synthesize evaluation practices, domain specific applications, and architectural variants such as Naive, Advanced, and Modular RAG. Finally, we outline persistent challenges related to retrieval quality, reliability, domain adaptation, scalability, and explainability, and identify opportunities for building RAG systems that are more reliable, adaptable, and transparent.

著者のコメント

published in Artificial intelligence reviews

arXiv ID: 2610.01936 / 要約の誤りについて