A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models
2024/05/10 by Fan, Wenqi, Ding, Yujuan, Ning, Liangbo +5 · 124 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Information Retrieval (cs.IR)
paper · doi:10.48550/arxiv.2405.06211
Abstract
As one of the most advanced techniques in AI, Retrieval-Augmented Generation (RAG) can offer reliable and up-to-date external knowledge, providing huge convenience for numerous tasks. Particularly in the era of AI-Generated Content (AIGC), the powerful capacity of retrieval in providing additional knowledge enables RAG to assist existing generative AI in producing high-quality outputs. Recently, Large Language Models (LLMs) have demonstrated revolutionary abilities in language understanding and generation, while still facing inherent limitations, such as hallucinations and out-of-date internal knowledge. Given the powerful abilities of RAG in providing the latest and helpful auxiliary information, Retrieval-Augmented Large Language Models (RA-LLMs) have emerged to harness external and authoritative knowledge bases, rather than solely relying on the model's internal knowledge, to augment the generation quality of LLMs. In this survey, we comprehensively review existing research studies in RA-LLMs, covering three primary technical perspectives: architectures, training strategies, and applications. As the preliminary knowledge, we briefly introduce the foundations and recent advances of LLMs. Then, to illustrate the practical significance of RAG for LLMs, we systematically review mainstream relevant work by their architectures, training strategies, and application areas, detailing specifically the challenges of each and the corresponding capabilities of RA-LLMs. Finally, to deliver deeper insights, we discuss current limitations and several promising directions for future research. Updated information about this survey can be found at https://advanced-recommender-systems.github.io/RAG-Meets-LLMs/
Cited by
- Problems With Large Language Models for Learner Modelling: Why LLMs Alone Fall Short for Responsible Tutoring in K--12 Education
- MPR-CiteG: Enhancing RAG with Multi-Portfolio Retrieval and Citation-Grounded Generation
- TAMEing Long Contexts in Personalization: Towards Training-Free and State-Aware MLLM Personalized Assistant
- X-GridAgent: An LLM-Powered Agentic AI System for Assisting Power Grid Analysis
- M3KG-RAG: Multi-hop Multimodal Knowledge Graph-enhanced Retrieval-Augmented Generation
- A Large-Language-Model Framework for Automated Humanitarian Situation Reporting
- CIRR: Causal-Invariant Retrieval-Augmented Recommendation with Faithful Explanations under Distribution Shift
- Graph-based Nearest Neighbors with Dynamic Updates via Random Walks
- Scalable Distributed Vector Search via Accuracy Preserving Index Construction
- A systematic assessment of Large Language Models for constructing two-level fractional factorial designs
- Plausibility as Failure: How LLMs and Humans Co-Construct Epistemic Error
- R4: Retrieval-Augmented Reasoning for Vision-Language Models in 4D Spatio-Temporal Space
- Leveraging Spreading Activation for Improved Document Retrieval in Knowledge-Graph-Based RAG Systems
- IaC Generation with LLMs: An Error Taxonomy and A Study on Configuration Knowledge Injection
- AIAuditTrack: A Framework for AI Security system
- Autonomous Construction-Site Safety Inspection Using Mobile Robots: A Multilayer VLM-LLM Pipeline
- From Context to EDUs: Faithful and Structured Context Compression via Elementary Discourse Unit Decomposition
- A Systematic Characterization of LLM Inference on GPUs
- SHRAG: AFrameworkfor Combining Human-Inspired Search with RAG
- HKRAG: Holistic Knowledge Retrieval-Augmented Generation Over Visually-Rich Documents
- HyperbolicRAG: Enhancing Retrieval-Augmented Generation with Hyperbolic Representations
- VecIntrinBench: Benchmarking Cross-Architecture Intrinsic Code Migration for RISC-V Vector
- The Oracle and The Prism: A Decoupled and Efficient Framework for Generative Recommendation Explanation
- WebRec: Enhancing LLM-based Recommendations with Attention-guided RAG from Web
- NeuroPath: Neurobiology-Inspired Path Tracking and Reflection for Semantically Coherent Retrieval
- Grounded by Experience: Generative Healthcare Prediction Augmented with Hierarchical Agentic Retrieval
- BudgetLeak: Membership Inference Attacks on RAG Systems via the Generation Budget Side Channel
- Structured RAG for Answering Aggregative Questions
- Rethinking Retrieval-Augmented Generation for Medicine: A Large-Scale, Systematic Expert Evaluation and Practical Insights
- QuAnTS: Question Answering on Time Series
- Beyond Single Embeddings: Capturing Diverse Targets with Multi-Query Retrieval
- AGRAG: Advanced Graph-based Retrieval-Augmented Generation for LLMs
- A Systematic Literature Review of Code Hallucinations in LLMs: Characterization, Mitigation Methods, Challenges, and Future Directions for Reliable AI
- TreeQA: Enhanced LLM-RAG with logic tree reasoning for reliable and interpretable multi-hop question answering
- Adapting Large Language Models to Emerging Cybersecurity using Retrieval Augmented Generation
- Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governance in Generative AI
- State of the Art of LLM-Enabled Interaction with Visualization
- ScaleCall -- Agentic Tool Calling at Scale for Fintech: Challenges, Methods, and Deployment Insights
- Metadata-Driven Retrieval-Augmented Generation for Financial Question Answering
- Graph-Guided Concept Selection for Efficient Retrieval-Augmented Generation
- M-Eval: A Heterogeneity-Based Framework for Multi-evidence Validation in Medical RAG Systems
- Dynamically Detect and Fix Hardness for Efficient Approximate Nearest Neighbor Search
- Foundation of Intelligence: Review of Math Word Problems from Human Cognition Perspective
- NeuroGenPoisoning: Neuron-Guided Attacks on Retrieval-Augmented Generation of LLM via Genetic Optimization of External Knowledge
- HA-RAG: Hotness-Aware RAG Acceleration via Mixed Precision and Data Placement
- ResearchGPT: Benchmarking and Training LLMs for End-to-End Computer Science Research Workflows
- FidelityGPT: Correcting Decompilation Distortions with Retrieval Augmented Generation
- Enhancing Hotel Recommendations with AI: LLM-Based Review Summarization and Query-Driven Insights
- AtlasKV: Augmenting LLMs with Billion-Scale Knowledge Graphs in 20GB VRAM
- Comprehending Spatio-temporal Data via Cinematic Storytelling using Large Language Models
- Efficient Toxicity Detection in Gaming Chats: A Comparative Study of Embeddings, Fine-Tuned Transformers and LLMs
- A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Evaluations, and Applications
- Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation
- Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning
- Query-Specific GNN: A Comprehensive Graph Representation Learning Method for Retrieval Augmented Generation
- Autonomous Agents for Scientific Discovery: Orchestrating Scientists, Language, Code, and Physics
- DualResearch: Entropy-Gated Dual-Graph Retrieval for Answer Reconstruction
- RCPU: Rotation-Constrained Error Compensation for Structured Pruning of Large Language Models
- Agentic generative AI for media content discovery at the national football league
- Exposing Citation Vulnerabilities in Generative Engines
- MARS: Co-evolving Dual-System Deep Research via Multi-Agent Reinforcement Learning
- Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches
- A Lightweight Large Language Model-Based Multi-Agent System for 2D Frame Structural Analysis
- UNIDOC-BENCH: A Unified Benchmark for Document-Centric Multimodal RAG
- External Data Extraction Attacks against Retrieval-Augmented Large Language Models
- Copy-Paste to Mitigate Large Language Model Hallucinations
- Attribution Gradients: Incrementally Unfolding Citations for Critical Examination of Attributed AI Answers
- MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent Environments
- AutoLabs: Cognitive Multi-Agent Systems with Self-Correction for Autonomous Chemical Experimentation
- Beyond Static Retrieval: Opportunities and Pitfalls of Iterative Retrieval in GraphRAG
- SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA
- SafeSearch: Automated Red-Teaming for the Safety of LLM-Based Search Agents
- LUMINA: Detecting Hallucinations in RAG System with Context-Knowledge Signals
- D-Artemis: A Deliberative Cognitive Framework for Mobile GUI Multi-Agents
- SGMem: Sentence Graph Memory for Long-Term Conversational Agents
- Hierarchical Reranking for Scalable Financial RAG System
- A Knowledge Graph-based Retrieval-Augmented Generation Framework for Algorithm Selection in the Facility Layout Problem
- Revealing Multimodal Causality with Large Language Models
- CORE-RAG: Lossless Compression for Retrieval-Augmented LLMs via Reinforcement Learning
- Towards the Distributed Large-scale k-NN Graph Construction by Graph Merge
- A Survey on Retrieval And Structuring Augmented Generation with Large Language Models
- Approximate Graph Propagation Revisited: Dynamic Parameterized Queries, Tighter Bounds and Dynamic Updates
- AgentX: Towards Orchestrating Robust Agentic Workflow Patterns with FaaS-hosted MCP Services
- Electricity Demand and Grid Impacts of AI Data Centers: Challenges and Prospects
- Chain or tree? Re-evaluating complex reasoning from the perspective of a matrix of thought
- Lighting the Way for BRIGHT: Reproducible Baselines with Anserini, Pyserini, and RankLLM
- Enhancing Reliability in LLM-Integrated Robotic Systems: A Unified Approach to Security and Safety
- RAG-PRISM: A Personalized, Rapid, and Immersive Skill Mastery Framework with Adaptive Retrieval-Augmented Tutoring
- MSRS: Evaluating Multi-Source Retrieval-Augmented Generation
- Model-Driven Quantum Code Generation Using Large Language Models and Retrieval-Augmented Generation
- LFD: Layer Fused Decoding to Exploit External Knowledge in Retrieval-Augmented Generation
- Diverse And Private Synthetic Datasets Generation for RAG evaluation: A multi-agent framework
- Explicit v.s. Implicit Memory: Exploring Multi-hop Complex Reasoning Over Personalized Information
- Atom-Searcher: Enhancing Agentic Deep Research via Fine-Grained Atomic Thought Reward
- Retrieval-augmented reasoning with lean language models
- SMA: Who Said That? Auditing Membership Leakage in Semi-Black-box RAG Controlling
- Understanding Users' Privacy Perceptions Towards LLM's RAG-based Memory
- Multi-Modal Requirements Data-based Acceptance Criteria Generation using LLMs
- Integrating Rules and Semantics for LLM-Based C-to-Rust Translation
- RAGTrace: Understanding and Refining Retrieval-Generation Dynamics in Retrieval-Augmented Generation
- mKG-RAG: Multimodal Knowledge Graph-Enhanced RAG for Visual Question Answering
- Retrieval-Augmented Water Level Forecasting for Everglades
- TURA: Tool-Augmented Unified Retrieval Agent for AI Search
- Method-Based Reasoning for Large Language Models: Extraction, Reuse, and Continuous Improvement
- TreeRanker: Fast and Model-agnostic Ranking System for Code Suggestions in IDEs
- Prompting Large Language Models with Partial Knowledge for Answering Questions with Unseen Entities
- Provably Secure Retrieval-Augmented Generation
- GraphRAG-R1: Graph Retrieval-Augmented Generation with Process-Constrained Reinforcement Learning
- AutoBridge: Automating Smart Device Integration with Centralized Platform
- Fast and Accurate Contextual Knowledge Extraction Using Cascading Language Model Chains and Candidate Answers
- Rote Learning Considered Useful: Generalizing over Memorized Data in LLMs
- Conversations over Clicks: Impact of Chatbots on Information Search in Interdisciplinary Learning
- Graph-Augmented Large Language Model Agents: Current Progress and Future Prospects
- Advancing Shared and Multi-Agent Autonomy in Underwater Missions: Integrating Knowledge Graphs and Retrieval-Augmented Generation
- GREAT: Guiding Query Generation with a Trie for Recommending Related Search about Video at Kuaishou
- CONCAP: Seeing Beyond English with Concepts Retrieval-Augmented Captioning
- OmniBench-RAG: A Multi-Domain Evaluation Platform for Retrieval-Augmented Generation Tools
- A Systematic Review of Key Retrieval-Augmented Generation (RAG) Systems: Progress, Gaps, and Future Directions
- Transform Before You Query: A Privacy-Preserving Approach for Vector Retrieval with Embedding Space Alignment
- A Comprehensive Review on Harnessing Large Language Models to Overcome Recommender System Challenges
- DyG-RAG: Dynamic Graph Retrieval-Augmented Generation with Event-Centric Reasoning
- Evaluating LLMs on Sequential API Call Through Automated Test Generation
- KGRAG-Ex: Explainable Retrieval-Augmented Generation with Knowledge Graph-based Perturbations
- Clue-RAG: Towards Accurate and Cost-Efficient Graph-based RAG via Multi-Partite Graph and Query-Driven Iterative Retrieval
Related