paper-with-me

Papers

Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents

2024-05-05 · Junkai Li, Yunghwei Lai, Weitao Li, Jingyi Ren, Meng Zhang, Xinhui Kang, Siyu Wang, Peng Li, Ya-Qin Zhang, Weizhi Ma, Yang Liu

The recent rapid development of large language models (LLMs) has sparked a new wave of technological revolution in medical artificial intelligence (AI). While LLMs are designed to understand and generate text like a human, autonomous agents that utilize LLMs as their "brain" have exhibited capabilities beyond text processing such as planning, reflection, and using tools by enabling their "bodies" to interact with the environment. We introduce a simulacrum of hospital called Agent Hospital that simulates the entire process of treating illness, in which all patients, nurses, and doctors are LLM-powered autonomous agents. Within the simulacrum, doctor agents are able to evolve by treating a large number of patient agents without the need to label training data manually. After treating tens of thousands of patient agents in the simulacrum (human doctors may take several years in the real world), the evolved doctor agents outperform state-of-the-art medical agent methods on the MedQA benchmark comprising US Medical Licensing Examination (USMLE) test questions. Our methods of simulacrum construction and agent evolution have the potential in benefiting a broad range of applications beyond medical AI.

📄 PDF Abstract BibTeX arXiv:2405.02957

Code (0)

등록된 구현이 없습니다.

Tasks

MedQAQuestion Answering

Similar Papers 제목 키워드 기반

LatentDx: Latent Multi-Agent Communication for Cross-Hospital Rare-Disease Diagnosis

2026-06-11 · Ziqing Wang, Lili Zhao, Kaize Ding arxiv

Rare diseases affect over $300$ million patients across more than $7{,}000$ conditions, yet no single hospital encounters enough cases of any one condition for reliable diagnosis. Cross-hospital collaboration could help …

The Application of MATEC (Multi-AI Agent Team Care) Framework in Sepsis Care

2025-02-09 · Andrew Cho, Jason M. Woo, Brian Shi, Aishwaryaa Udeshi 외

Under-resourced or rural hospitals have limited access to medical specialists and healthcare professionals, which can negatively impact patient outcomes in sepsis. To address this gap, we developed the MATEC (Multi-AI Ag…

AI Agent

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence

2026-03-16 · Peigen Liu, Rui Ding, Yuren Mao, Ziyan Jiang 외 arxiv

Large Language Model (LLM)-based Collective Intelligence (CI) presents a promising approach to overcoming the data wall and continuously boosting the capabilities of LLM agents. However, there is currently no dedicated a…

CP-Env: Evaluating Large Language Models on Clinical Pathways in a Controllable Hospital Environment

2025-12-11 · Yakun Zhu, Zhongzhen Huang, Qianhan Feng, Linjie Mu 외 arxiv

Medical care follows complex clinical pathways that extend beyond isolated physician-patient encounters, emphasizing decision-making and transitions between different stages. Current benchmarks focusing on static exams o…

AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

2024-02-15 · Zhihao Fan, Jialong Tang, Wei Chen, Siyuan Wang 외

Artificial intelligence has significantly advanced healthcare, particularly through large language models (LLMs) that excel in medical question answering benchmarks. However, their real-world clinical application remains…

BenchmarkingDiagnosticMedical Question AnsweringQuestion Answering