Improving large language model applications in biomedicine with retrieval-augmented generation: a systematic review, meta-analysis, and clinical development guidelines

被引:1
|
作者
Liu, Siru [1 ,2 ]
Mccoy, Allison B. [1 ]
Wright, Adam [1 ,3 ]
机构
[1] Vanderbilt Univ, Med Ctr, Dept Biomed Informat, 2525 West End Ave 1475, Nashville, TN 37212 USA
[2] Vanderbilt Univ, Dept Comp Sci, Nashville, TN 37212 USA
[3] Vanderbilt Univ, Med Ctr, Dept Med, Nashville, TN 37212 USA
基金
美国国家卫生研究院;
关键词
large language model; retrieval augmented generation; systematic review; meta-analysis; BIAS;
D O I
10.1093/jamia/ocaf008
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Objective The objectives of this study are to synthesize findings from recent research of retrieval-augmented generation (RAG) and large language models (LLMs) in biomedicine and provide clinical development guidelines to improve effectiveness.Materials and Methods We conducted a systematic literature review and a meta-analysis. The report was created in adherence to the Preferred Reporting Items for Systematic Reviews and Meta-Analyses 2020 analysis. Searches were performed in 3 databases (PubMed, Embase, PsycINFO) using terms related to "retrieval augmented generation" and "large language model," for articles published in 2023 and 2024. We selected studies that compared baseline LLM performance with RAG performance. We developed a random-effect meta-analysis model, using odds ratio as the effect size.Results Among 335 studies, 20 were included in this literature review. The pooled effect size was 1.35, with a 95% confidence interval of 1.19-1.53, indicating a statistically significant effect (P = .001). We reported clinical tasks, baseline LLMs, retrieval sources and strategies, as well as evaluation methods.Discussion Building on our literature review, we developed Guidelines for Unified Implementation and Development of Enhanced LLM Applications with RAG in Clinical Settings to inform clinical applications using RAG.Conclusion Overall, RAG implementation showed a 1.35 odds ratio increase in performance compared to baseline LLMs. Future research should focus on (1) system-level enhancement: the combination of RAG and agent, (2) knowledge-level enhancement: deep integration of knowledge into LLM, and (3) integration-level enhancement: integrating RAG systems within electronic health records.
引用
收藏
页数:11
相关论文
共 50 条
  • [1] Retrieval-Augmented Generation Approach: Document Question Answering using Large Language Model
    Muludi, Kurnia
    Fitria, Kaira Milani
    Triloka, Joko
    Sutedi
    INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2024, 15 (03) : 776 - 785
  • [2] Injury degree appraisal of large language model based on retrieval-augmented generation and deep learning
    Zhang, Fan
    Luo, Yifang
    Gao, Zihuan
    Han, Aihua
    INTERNATIONAL JOURNAL OF LAW AND PSYCHIATRY, 2025, 100
  • [3] Application of retrieval-augmented generation for interactive industrial knowledge management via a large language model
    Chen, Lun-Chi
    Pardeshi, Mayuresh Sunil
    Liao, Yi-Xiang
    Pai, Kai-Chih
    COMPUTER STANDARDS & INTERFACES, 2025, 94
  • [4] Evaluation of the integration of retrieval-augmented generation in large language model for breast cancer nursing care responses
    Xu, Ruiyu
    Hong, Ying
    Zhang, Feifei
    Xu, Hongmei
    SCIENTIFIC REPORTS, 2024, 14 (01):
  • [5] Retrieval-Augmented Generation-aided causal identification of aviation accidents: A large language model methodology
    Ren, Tengfei
    Zhang, Zhipeng
    Jia, Bo
    Zhang, Shiwen
    EXPERT SYSTEMS WITH APPLICATIONS, 2025, 278
  • [6] Systematic review and meta-analysis of cardiac neurosis for development of clinical practice guidelines of Korean medicine
    Park, Hui-Yeong
    Lee, Hyun Woo
    Song, Geum-Ju
    Hong, Sunggyu
    Hong, Sunghee
    Suh, Hyo-Weon
    Yoon, Seok-In
    Park, Chan
    Chung, Sun-Yong
    Kim, Jong Woo
    FRONTIERS IN PSYCHIATRY, 2024, 15
  • [7] Fine-Tuning Retrieval-Augmented Generation with an Auto-Regressive Language Model for Sentiment Analysis in Financial Reviews
    Mathebula, Miehleketo
    Modupe, Abiodun
    Marivate, Vukosi
    APPLIED SCIENCES-BASEL, 2024, 14 (23):
  • [8] Clinical practice guidelines on polycystic ovary syndrome: a systematic review and comparative meta-analysis
    Dun, J.
    Wang, X.
    Yang, J.
    Xu, J.
    CLINICAL AND EXPERIMENTAL OBSTETRICS & GYNECOLOGY, 2020, 47 (04): : 465 - 471
  • [9] Clinical applications of contactless photoplethysmography for monitoring in adults: A systematic review and meta-analysis
    Bautista, Melissa Joanne
    Kowal, Mikolaj
    Cave, Daniel G. W.
    Downey, Candice
    Jayne, David G.
    JOURNAL OF CLINICAL AND TRANSLATIONAL SCIENCE, 2023, 7 (01)
  • [10] Applications and effectiveness of augmented reality in safety training: A systematic literature review and meta-analysis
    Gong, Peizhen
    Lu, Ying
    Lovreglio, Ruggiero
    Lv, Xiaofeng
    Chi, Zexun
    SAFETY SCIENCE, 2024, 178