paper-with-me

Papers

Streamlining Systematic Reviews: A Novel Application of Large Language Models

2024-12-14 · Fouad Trad, Ryan Yammine, Jana Charafeddine, Marlene Chakhtoura, Maya Rahme, Ghada El-Hajj Fuleihan, Ali Chehab

Systematic reviews (SRs) are essential for evidence-based guidelines but are often limited by the time-consuming nature of literature screening. We propose and evaluate an in-house system based on Large Language Models (LLMs) for automating both title/abstract and full-text screening, addressing a critical gap in the literature. Using a completed SR on Vitamin D and falls (14,439 articles), the LLM-based system employed prompt engineering for title/abstract screening and Retrieval-Augmented Generation (RAG) for full-text screening. The system achieved an article exclusion rate (AER) of 99.5%, specificity of 99.6%, a false negative rate (FNR) of 0%, and a negative predictive value (NPV) of 100%. After screening, only 78 articles required manual review, including all 20 identified by traditional methods, reducing manual screening time by 95.5%. For comparison, Rayyan, a commercial tool for title/abstract screening, achieved an AER of 72.1% and FNR of 5% when including articles Rayyan considered as undecided or likely to include. Lowering Rayyan's inclusion thresholds improved FNR to 0% but increased screening time. By addressing both screening phases, the LLM-based system significantly outperformed Rayyan and traditional methods, reducing total screening time to 25.5 hours while maintaining high accuracy. These findings highlight the transformative potential of LLMs in SR workflows by offering a scalable, efficient, and accurate solution, particularly for the full-text screening phase, which has lacked automation tools.

📄 PDF Abstract BibTeX arXiv:2412.15247

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesPrompt EngineeringRAGRetrieval-augmented GenerationSpecificity

Similar Papers 제목 키워드 기반

Accelerating Clinical Evidence Synthesis with Large Language Models

2024-06-25 · Zifeng Wang, Lang Cao, Benjamin Danek, Qiao Jin 외

Synthesizing clinical evidence largely relies on systematic reviews of clinical trials and retrospective analyses from medical literature. However, the rapid expansion of publications presents challenges in efficiently i…

Language Modelling

Automating Research Synthesis with Domain-Specific Large Language Model Fine-Tuning

2024-04-08 · Teo Susnjak, Peter Hwang, Napoleon H. Reyes, Andre L. C. Barczak 외

This research pioneers the use of fine-tuned Large Language Models (LLMs) to automate Systematic Literature Reviews (SLRs), presenting a significant and novel contribution in integrating AI to enhance academic research m…

HallucinationLanguage ModelingLanguage ModellingLarge Language Model

Streamlining the Selection Phase of Systematic Literature Reviews (SLRs) Using AI-Enabled GPT-4 Assistant API

2024-01-14 · Seyed Mohammad Ali Jafari

The escalating volume of academic literature presents a formidable challenge in staying updated with the newest research developments. Addressing this, this study introduces a pioneering AI-based tool, configured specifi…

Management

Enhancing Systematic Reviews with Large Language Models: Using GPT-4 and Kimi

2025-04-28 · Dandan Chen Kaptur, Yue Huang, Xuejun Ryan Ji, Yanhui Guo 외

This research delved into GPT-4 and Kimi, two Large Language Models (LLMs), for systematic reviews. We evaluated their performance by comparing LLM-generated codes with human-generated codes from a peer-reviewed systemat…

A foundation model for human-AI collaboration in medical literature mining

2025-01-27 · Zifeng Wang, Lang Cao, Qiao Jin, Joey Chan 외

Systematic literature review is essential for evidence-based medicine, requiring comprehensive analysis of clinical trial publications. However, the application of artificial intelligence (AI) models for medical literatu…

Literature MiningSystematic Literature Review