paper-with-me

홈 › Papers

Evaluation of AI Chatbots for Patient-Specific EHR Questions

2023-06-05 · Alaleh Hamidi, Kirk Roberts

This paper investigates the use of artificial intelligence chatbots for patient-specific question answering (QA) from clinical notes using several large language model (LLM) based systems: ChatGPT (versions 3.5 and 4), Google Bard, and Claude. We evaluate the accuracy, relevance, comprehensiveness, and coherence of the answers generated by each model using a 5-point Likert scale on a set of patient-specific questions.

📄 PDF Abstract BibTeX arXiv:2306.02549

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language ModelQuestion Answering

Similar Papers 제목 키워드 기반

Large language models provide unsafe answers to patient-posed medical questions

2025-07-25 · Rachel L. Draelos, Samina Afreen, Barbara Blasko, Tiffany L. Brazile 외 arxiv

Millions of patients are already using large language model (LLM) chatbots for medical advice on a regular basis, raising patient safety concerns. This physician-led red-teaming study compares the safety of four publicly…

LLM-empowered Chatbots for Psychiatrist and Patient Simulation: Application and Evaluation

2023-05-23 · Siyuan Chen, Mengyue Wu, Kenny Q. Zhu, Kunyao Lan 외

Empowering chatbots in the field of mental health is receiving increasing amount of attention, while there still lacks exploration in developing and evaluating chatbots in psychiatric outpatient scenarios. In this work, …

ChatbotDiagnostic

Fine-tuning Large Language Model (LLM) Artificial Intelligence Chatbots in Ophthalmology and LLM-based evaluation using GPT-4

2024-02-15 · Ting Fang Tan, Kabilan Elangovan, Liyuan Jin, Yao Jie 외

Purpose: To assess the alignment of GPT-4-based evaluation to human clinician experts, for the evaluation of responses to ophthalmology-related patient queries generated by fine-tuned LLM chatbots. Methods: 400 ophthalmo…

ChatbotLanguage ModelingLanguage ModellingLarge Language Model

Building Chatbots from Forum Data: Model Selection Using Question Answering Metrics

2017-10-02 · RANLP 2017 9 · Martin Boyanov, Ivan Koychev, Preslav Nakov, Alessandro Moschitti 외

We propose to use question answering (QA) data from Web forums to train chatbots from scratch, i.e., without dialog training data. First, we extract pairs of question and answer sentences from the typically much longer t…

Model SelectionQuestion Answering

Do AI chatbots find what experts would? Effects of model, user role, and sample size on study retrieval for medical questions

2026-08-13 · Qingfang Liu, Qiao Jin, Joe D. Menke, Thorsten Kahnt 외 arxiv

Large language model (LLM) chatbots are increasingly used to answer clinical questions with citations to relevant clinical studies. Prior research has largely focused on citation fabrication, leaving a gap in evaluating …