paper-with-me

Papers

Evaluating the Application of ChatGPT in Outpatient Triage Guidance: A Comparative Study

2024-04-27 · Dou Liu, Ying Han, Xiandi Wang, Xiaomei Tan, Di Liu, Guangwu Qian, Kang Li, Dan Pu, Rong Yin

The integration of Artificial Intelligence (AI) in healthcare presents a transformative potential for enhancing operational efficiency and health outcomes. Large Language Models (LLMs), such as ChatGPT, have shown their capabilities in supporting medical decision-making. Embedding LLMs in medical systems is becoming a promising trend in healthcare development. The potential of ChatGPT to address the triage problem in emergency departments has been examined, while few studies have explored its application in outpatient departments. With a focus on streamlining workflows and enhancing efficiency for outpatient triage, this study specifically aims to evaluate the consistency of responses provided by ChatGPT in outpatient guidance, including both within-version response analysis and between-version comparisons. For within-version, the results indicate that the internal response consistency for ChatGPT-4.0 is significantly higher than ChatGPT-3.5 (p=0.03) and both have a moderate consistency (71.2% for 4.0 and 59.6% for 3.5) in their top recommendation. However, the between-version consistency is relatively low (mean consistency score=1.43/3, median=1), indicating few recommendations match between the two versions. Also, only 50% top recommendations match perfectly in the comparisons. Interestingly, ChatGPT-3.5 responses are more likely to be complete than those from ChatGPT-4.0 (p=0.02), suggesting possible differences in information processing and response generation between the two versions. The findings offer insights into AI-assisted outpatient operations, while also facilitating the exploration of potentials and limitations of LLMs in healthcare utilization. Future research may focus on carefully optimizing LLMs and AI integration in healthcare systems based on ergonomic and human factors principles, precisely aligning with the specific needs of effective outpatient triage.

📄 PDF Abstract BibTeX arXiv:2405.00728

Code (0)

등록된 구현이 없습니다.

Tasks

Response Generation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Medical Triage as Pairwise Ranking: A Benchmark for Urgency in Patient Portal Messages

2026-01-19 · Joseph Gatto, Parker Seegmiller, Timothy Burdick, Philip Resnik 외 arxiv

Medical triage is the task of allocating medical resources and prioritizing patients based on medical need. This paper introduces the first large-scale public dataset for studying medical triage in the context of asynchr…

LLM-empowered Chatbots for Psychiatrist and Patient Simulation: Application and Evaluation

2023-05-23 · Siyuan Chen, Mengyue Wu, Kenny Q. Zhu, Kunyao Lan 외

Empowering chatbots in the field of mental health is receiving increasing amount of attention, while there still lacks exploration in developing and evaluating chatbots in psychiatric outpatient scenarios. In this work, …

ChatbotDiagnostic

Dental-TriageBench: Benchmarking Multimodal Reasoning for Hierarchical Dental Triage

2026-03-18 · Ziyi He, Yushi Feng, Shuangyu Yang, Yinghao Zhu 외 arxiv

Dental triage is a safety-critical clinical routing task that requires integrating multimodal clinical information (e.g., patient complaints and radiographic evidence) to determine complete referral plans. We present Den…

Multimodal Reasoning

Large Language Models for Outpatient Referral: Problem Definition, Benchmarking and Challenges

2025-03-11 · Xiaoxiao Liu, Qingying Xiao, Junying Chen, Xiangyi Feng 외

Large language models (LLMs) are increasingly applied to outpatient referral tasks across healthcare systems. However, there is a lack of standardized evaluation criteria to assess their effectiveness, particularly in dy…

Benchmarking

Gender-Dependent Diagnostic Substitution in LLM Medical Triage: Same Symptoms, Unequal Urgency

2026-06-02 · Qi Han Wong arxiv

We investigate whether large language models produce different medical triage recommendations for identical neurological symptoms when only the patient's stated gender and age vary. Using three model families--Gemini 3.5…