paper-with-me

홈 › Papers

MedClarify: An information-seeking AI agent for medical diagnosis with case-specific follow-up questions

2026-02-19 · Hui Min Wong, Philip Heesen, Pascal Janetzky, Martin Bendszus, Stefan Feuerriegel arxiv

Large language models (LLMs) are increasingly used for diagnostic tasks in medicine. In clinical practice, the correct diagnosis can rarely be immediately inferred from the initial patient presentation alone. Rather, reaching a diagnosis often involves systematic history taking, during which clinicians reason over multiple potential conditions through iterative questioning to resolve uncertainty. This process requires considering differential diagnoses and actively excluding emergencies that demand immediate intervention. Yet, the ability of medical LLMs to generate informative follow-up questions and thus reason over differential diagnoses remains underexplored. Here, we introduce MedClarify, an AI agent for information-seeking that can generate follow-up questions for iterative reasoning to support diagnostic decision-making. Specifically, MedClarify computes a list of candidate diagnoses analogous to a differential diagnosis, and then proactively generates follow-up questions aimed at reducing diagnostic uncertainty. By selecting the question with the highest expected information gain, MedClarify enables targeted, uncertainty-aware reasoning to improve diagnostic performance. In our experiments, we first demonstrate the limitations of current LLMs in medical reasoning, which often yield multiple, similarly likely diagnoses, especially when patient cases are incomplete or relevant information for diagnosis is missing. We then show that our information-theoretic reasoning approach can generate effective follow-up questioning and thereby reduces diagnostic errors by ~27 percentage points (p.p.) compared to a standard single-shot LLM baseline. Altogether, MedClarify offers a path to improve medical LLMs through agentic information-seeking and to thus promote effective dialogues with medical LLMs that reflect the iterative and uncertain nature of real-world clinical reasoning.

📄 PDF Abstract BibTeX arXiv:2602.17308

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Diagnosis

Similar Papers 제목 키워드 기반

Information-seeking failures of large language models in agentic clinical reasoning

2026-07-11 · Krischan Braitsch, Laura K. Schmalbrock, Theresa Weltermann, Andrew F. Berdel 외 arxiv

Large language models achieve high scores on medical knowledge assessments, yet clinical reasoning requires actively deciding what to investigate under uncertainty. We developed an agentic evaluation framework in hematol…

PathFound: An Agentic Multimodal Model Activating Evidence-seeking Pathological Diagnosis

2025-12-29 · Shengyi Hua, Jianfeng Wu, Tianle Shen, Kangzhe Hu 외 arxiv

Recent pathological foundation models have substantially advanced visual representation learning and multimodal interaction. However, most models still rely on a static inference paradigm in which whole-slide images are …

Representation LearningReinforcement Learning

Shoot First, Ask Questions Later? Building Rational Agents that Explore and Act Like People

2025-10-23 · Gabriel Grand, Valerio Pepe, Jacob Andreas, Joshua B. Tenenbaum arxiv

Many emerging applications of AI--from scientific discovery to medical diagnosis--require agents to seek information strategically: forming hypotheses, asking targeted questions, and making decisions under uncertainty. I…

Medical Diagnosis

3MDBench: Medical Multimodal Multi-agent Dialogue Benchmark

2025-03-26 · Ivan Sviridov, Amina Miftakhova, Artemiy Tereshchenko, Galina Zubkova 외

Though Large Vision-Language Models (LVLMs) are being actively explored in medicine, their ability to conduct telemedicine consultations combining accurate diagnosis with professional dialogue remains underexplored. In t…

DiagnosticMultimodal Reasoning

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning

2026-05-19 · Juncheng Wu, Letian Zhang, Yuhan Wang, Haoqin Tu 외 arxiv

Large language models (LLMs) and agentic systems have shown promise for clinical decision support, but existing works largely assume that evidence has already been curated and handed to the model. Real-world clinical wor…