paper-with-me

Papers

Almanac: Retrieval-Augmented Language Models for Clinical Medicine

2023-03-01 · Cyril Zakka, Akash Chaurasia, Rohan Shad, Alex R. Dalal, Jennifer L. Kim, Michael Moor, Kevin Alexander, Euan Ashley, Jack Boyd, Kathleen Boyd, Karen Hirsch, Curt Langlotz, Joanna Nelson, William Hiesinger

Large-language models have recently demonstrated impressive zero-shot capabilities in a variety of natural language tasks such as summarization, dialogue generation, and question-answering. Despite many promising applications in clinical medicine, adoption of these models in real-world settings has been largely limited by their tendency to generate incorrect and sometimes even toxic statements. In this study, we develop Almanac, a large language model framework augmented with retrieval capabilities for medical guideline and treatment recommendations. Performance on a novel dataset of clinical scenarios (n = 130) evaluated by a panel of 5 board-certified and resident physicians demonstrates significant increases in factuality (mean of 18% at p-value < 0.05) across all specialties, with improvements in completeness and safety. Our results demonstrate the potential for large language models to be effective tools in the clinical decision-making process, while also emphasizing the importance of careful testing and deployment to mitigate their shortcomings.

📄 PDF Abstract BibTeX arXiv:2303.01229

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingDialogue GenerationLanguage ModelingLanguage ModellingLarge Language ModelQuestion AnsweringRetrieval

Similar Papers 제목 키워드 기반

Almanac Copilot: Towards Autonomous Electronic Health Record Navigation

2024-04-30 · Cyril Zakka, Joseph Cho, Gracia Fahed, Rohan Shad 외

Clinicians spend large amounts of time on clinical documentation, and inefficiencies impact quality of care and increase clinician burnout. Despite the promise of electronic medical records (EMR), the transition from pap…

Information RetrievalRetrieval

Retrieval-Augmented Generation in Medicine: A Scoping Review of Technical Implementations, Clinical Applications, and Ethical Considerations

2025-11-08 · Rui Yang, Matthew Yu Heng Wong, Huitao Li, Xin Li 외 arxiv

The rapid growth of medical knowledge and increasing complexity of clinical practice pose challenges. In this context, large language models (LLMs) have demonstrated value; however, inherent limitations remain. Retrieval…

Information ExtractionText SummarizationQuestion Answering

Lab-AI: Using Retrieval Augmentation to Enhance Language Models for Personalized Lab Test Interpretation in Clinical Medicine

2024-09-16 · Xiaoyu Wang, Haoyong Ouyang, Balu Bhasuran, Xiao Luo 외

Accurate interpretation of lab results is crucial in clinical medicine, yet most patient portals use universal normal ranges, ignoring conditional factors like age and gender. This study introduces Lab-AI, an interactive…

Language ModelingLanguage ModellingRAGRetrieval+1

ALMANACS: A Simulatability Benchmark for Language Model Explainability

2023-12-20 · Edmund Mills, Shiye Su, Stuart Russell, Scott Emmons

How do we measure the efficacy of language model explainability methods? While many explainability methods have been developed, they are typically evaluated on bespoke tasks, preventing an apples-to-apples comparison. To…

counterfactualLanguage ModelingLanguage Modellingmodel

Retrieval-Augmented Generation in Biomedicine: A Survey of Technologies, Datasets, and Clinical Applications

2025-05-02 · JiaWei He, Boya Zhang, Hossein Rouhizadeh, Yingjian Chen 외

Recent advances in large language models (LLMs) have demonstrated remarkable capabilities in natural language processing tasks. However, their application in the biomedical domain presents unique challenges, particularly…

RAGRetrievalRetrieval-augmented Generation