paper-with-me

Papers

A Study of Large Language Models for Patient Information Extraction: Model Architecture, Fine-Tuning Strategy, and Multi-task Instruction Tuning

2025-09-05 · Cheng Peng, Xinyu Dong, Mengxian Lyu, Daniel Paredes, Yaoyun Zhang, Yonghui Wu arxiv

Natural language processing (NLP) is a key technology to extract important patient information from clinical narratives to support healthcare applications. The rapid development of large language models (LLMs) has revolutionized many NLP tasks in the clinical domain, yet their optimal use in patient information extraction tasks requires further exploration. This study examines LLMs' effectiveness in patient information extraction, focusing on LLM architectures, fine-tuning strategies, and multi-task instruction tuning techniques for developing robust and generalizable patient information extraction systems. This study aims to explore key concepts of using LLMs for clinical concept and relation extraction tasks, including: (1) encoder-only or decoder-only LLMs, (2) prompt-based parameter-efficient fine-tuning (PEFT) algorithms, and (3) multi-task instruction tuning on few-shot learning performance. We benchmarked a suite of LLMs, including encoder-based LLMs (BERT, GatorTron) and decoder-based LLMs (GatorTronGPT, Llama 3.1, GatorTronLlama), across five datasets. We compared traditional full-size fine-tuning and prompt-based PEFT. We explored a multi-task instruction tuning framework that combines both tasks across four datasets to evaluate the zero-shot and few-shot learning performance using the leave-one-dataset-out strategy.

📄 PDF Abstract BibTeX arXiv:2509.04753

Code (0)

등록된 구현이 없습니다.

Tasks

parameter-efficient fine-tuningInformation ExtractionRelation ExtractionFew-Shot Learning

Similar Papers 제목 키워드 기반

PVminerLLM: Structured Extraction of Patient Voice from Patient-Generated Text using Large Language Models

2026-03-06 · Samah Fodeh, Linhai Ma, Ganesh Puthiaraju, Srivani Talakokkul 외 arxiv

Motivation: Patient-generated text contains critical information about patients' lived experiences, social circumstances, and engagement in care, including factors that strongly influence adherence, care coordination, an…

Extraction of Sleep Information from Clinical Notes of Patients with Alzheimer's Disease Using Natural Language Processing

2022-03-08 · Sonish Sivarajkumar, Thomas Yu CHow Tam, Haneef Ahamed Mohammad, Samual Viggiano 외

Alzheimer's Disease (AD) is the most common form of dementia in the United States. Sleep is one of the lifestyle-related factors that has been shown critical for optimal cognitive function in old age. However, there is a…

Language ModellingLarge Language ModelSleep Quality

Large Language Model-based Role-Playing for Personalized Medical Jargon Extraction

2024-08-10 · Jung Hoon Lim, Sunjae Kwon, Zonghai Yao, John P. Lalor 외

Previous studies reveal that Electronic Health Records (EHR), which have been widely adopted in the U.S. to allow patients to access their personal medical information, do not have high readability to patients due to the…

In-Context LearningLanguage ModelingLanguage ModellingLarge Language Model+1

An Information Extraction Approach to Prescreen Heart Failure Patients for Clinical Trials

2016-09-06 · Abhishek Kalyan Adupa, Ravi Prakash Garg, Jessica Corona-Cox, Sanjiv. J. Shah 외

To reduce the large amount of time spent screening, identifying, and recruiting patients into clinical trials, we need prescreening systems that are able to automate the data extraction and decision-making tasks that are…

Decision Making

Unsupervised extraction, labelling and clustering of segments from clinical notes

2022-11-21 · Petr Zelina, Jana Halámková, Vít Nováček

This work is motivated by the scarcity of tools for accurate, unsupervised information extraction from unstructured clinical notes in computationally underrepresented languages, such as Czech. We introduce a stepping sto…

Clustering