Skip to main navigation Skip to search Skip to main content

Retrieval-Augmented Generation for Natural Language Processing: A Survey

Research output: Journal Publications and ReviewsRGC 21 - Publication in refereed journalpeer-review

Abstract

Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that store knowledge. However, LLMs still suffer from several key issues, such as hallucination problems, knowledge update issues, and lacking domain-specific expertise. The appearance of retrieval-augmented generation (RAG), which leverages an external knowledge base to augment LLMs, mitigates these limitations. This paper presents a systematic review of RAG techniques for natural language processing (NLP), with a focus on retrievers and retrieval fusions. We introduce a novel taxonomy of retrieval fusions, such as query-based, logits-based, latent, and parametric fusion, and provide structured comparisons across accessibility, efficiency, and use cases. The paper further examines RAG applications across diverse NLP tasks, discusses evaluation methodologies and benchmark limitations, and analyzes training paradigms with and without knowledge base updates. Finally, we explore industrial deployment considerations and identify emerging challenges and future directions, including security, efficiency, and graph-based retrieval. © The Author(s) 2026.
Original languageEnglish
JournalArtificial Intelligence Review
Online published1 Jun 2026
DOIs
Publication statusOnline published - 1 Jun 2026

Bibliographical note

Research Unit(s) information for this publication is provided by the author(s) concerned.

Research Keywords

  • Retrieval-augmented generation
  • Natural language processing
  • Vector database
  • Large language model

Fingerprint

Dive into the research topics of 'Retrieval-Augmented Generation for Natural Language Processing: A Survey'. Together they form a unique fingerprint.

Cite this