Retrieval-augmented generation in medicine: A scoping review of technical implementations, clinical applications, and ethical considerations.
review · Level V
Where this comes from
- Record sourced from PubMed, PMID 42476143.
- Also identified by DOI 10.1016/j.xcrm.2026.102927.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
The rapid growth of medical knowledge and the increasing complexity of clinical practice pose challenges. In this context, large language models (LLMs) demonstrate value; however, inherent limitations remain. Retrieval-augmented generation (RAG) shows potential to enhance their clinical applicability. This study reviews RAG applications in medicine. We find that research primarily relies on publicly available data, with limited use of private data. For retrieval, approaches commonly rely on English-centric embedding models, while LLMs are mostly generic, with limited use of medical-specific LLMs. For evaluation, automated metrics evaluate generation quality and task performance, whereas human evaluation focuses on accuracy, completeness, relevance, and fluency, with insufficient attention to bias and safety. RAG applications are concentrated on question answering, report generation, text summarization, and information extraction. Overall, medical RAG remains at an early stage, requiring advances in clinical validation, cross-linguistic adaptation, and support for low-resource settings to enable trustworthy and responsible global use.