paper-with-me

홈 › Papers

How Much Would a Clinician Edit This Draft? Evaluating LLM Alignment for Patient Message Response Drafting

2026-01-16 · Parker Seegmiller, Joseph Gatto, Sarah E. Greer, Ganza Belise Isingizwe, Rohan Ray, Timothy E. Burdick, Sarah Masud Preum arxiv

Large language models (LLMs) show promise in drafting responses to patient portal messages, yet their integration into clinical workflows raises various concerns, including whether they would actually save clinicians time and effort in their portal workload. We investigate LLM alignment with individual clinicians through a comprehensive evaluation of the patient message response drafting task. We develop a novel taxonomy of thematic elements in clinician responses and propose a novel evaluation framework for assessing clinician editing load of LLM-drafted responses at both content and theme levels. We release an expert-annotated dataset and conduct large-scale evaluations of local and commercial LLMs using various adaptation techniques including thematic prompting, retrieval-augmented generation, supervised fine-tuning, and direct preference optimization. Our results reveal substantial epistemic uncertainty in aligning LLM drafts with clinician responses. While LLMs demonstrate capability in drafting certain thematic elements, they struggle with clinician-aligned generation in other themes, particularly question asking to elicit further information from patients. Theme-driven adaptation strategies yield improvements across most themes. Our findings underscore the necessity of adapting LLMs to individual clinician preferences to enable reliable and responsible use in patient-clinician communication workflows.

📄 PDF Abstract BibTeX arXiv:2601.11344

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Consumer-to-Clinical Language Shifts in Ambient AI Draft Notes and Clinician-Finalized Documentation: A Multi-level Analysis

2026-03-18 · Ha Na Cho, Yawen Guo, Sairam Sutari, Emilie Chow 외 arxiv

Ambient AI generates draft clinical notes from patient-clinician conversations, often using lay or consumer-oriented phrasing to support patient understanding instead of standardized clinical terminology. How clinicians …

Understanding Stigmatizing Language in Clinical Documentation: A Paired Comparison of Ambient AI Drafts and Clinician Finalized Notes

2026-04-14 · Yiliang Zhou, Yawen Guo, Sairam Sutari, Jasmine Dhillon 외 arxiv

Ambient artificial intelligence (AI) documentation tools are increasingly deployed to reduce clinician documentation burden, but their implications for biased language in clinical notes remain unclear. We conducted a lar…

Examine Clinicians' Modification of Hedging Language in Ambient AI Documentation: A Comparative Study of AI Drafts and Final Notes

2026-04-14 · Yiliang Zhou, Yawen Guo, Di Hu, Sairam Sutari 외 arxiv

Ambient AI documentation systems generate clinical note drafts that clinicians frequently revise before signing off into electronic health records, yet how these edits alter hedging language remains unclear. We conducted…

The impact of responding to patient messages with large language model assistance

2023-10-26 · Shan Chen, Marco Guevara, Shalini Moningi, Frank Hoebers 외

Documentation burden is a major contributor to clinician burnout, which is rising nationally and is an urgent threat to our ability to care for patients. Artificial intelligence (AI) chatbots, such as ChatGPT, could redu…

ChatbotDecision MakingLanguage ModelingLanguage Modelling+1

Human-in-the-Loop Interactive Report Generation for Chronic Disease Adherence

2026-01-10 · Xiaotian Zhang, Jinhong Yu, Pengwei Yan, Le Jiang 외 arxiv

Chronic disease management requires regular adherence feedback to prevent avoidable hospitalizations, yet clinicians lack time to produce personalized patient communications. Manual authoring preserves clinical accuracy …