paper-with-me

Papers

Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine

2024-11-20 · Yifan Yang, Qiao Jin, Robert Leaman, Xiaoyu Liu, Guangzhi Xiong, Maame Sarfo-Gyamfi, Changlin Gong, Santiago Ferrière-Steinert, W. John Wilbur, Xiaojun Li, Jiaxin Yuan, Bang An, Kelvin S. Castro, Francisco Erramuspe Álvarez, Matías Stockle, Aidong Zhang, Furong Huang, Zhiyong Lu

The remarkable capabilities of Large Language Models (LLMs) make them increasingly compelling for adoption in real-world healthcare applications. However, the risks associated with using LLMs in medical applications have not been systematically characterized. We propose using five key principles for safe and trustworthy medical AI: Truthfulness, Resilience, Fairness, Robustness, and Privacy, along with ten specific aspects. Under this comprehensive framework, we introduce a novel MedGuard benchmark with 1,000 expert-verified questions. Our evaluation of 11 commonly used LLMs shows that the current language models, regardless of their safety alignment mechanisms, generally perform poorly on most of our benchmarks, particularly when compared to the high performance of human physicians. Despite recent reports indicate that advanced LLMs like ChatGPT can match or even exceed human performance in various medical tasks, this study underscores a significant safety gap, highlighting the crucial need for human oversight and the implementation of AI safety guardrails.

📄 PDF Abstract BibTeX arXiv:2411.14487

Code (0)

등록된 구현이 없습니다.

Tasks

FairnessSafety Alignment

Similar Papers 제목 키워드 기반

Unveiling Trust in Multimodal Large Language Models: Evaluation, Analysis, and Mitigation

2025-08-21 · Yichi Zhang, Yao Huang, Yifan Wang, Yitong Sun 외 arxiv

The trustworthiness of Multimodal Large Language Models (MLLMs) remains an intense concern despite the significant progress in their capabilities. Existing evaluation and mitigation approaches often focus on narrow aspec…

TrustLLM: Trustworthiness in Large Language Models

2024-01-10 · Yue Huang, Lichao Sun, Haoran Wang, Siyuan Wu 외

Large language models (LLMs), exemplified by ChatGPT, have gained considerable attention for their excellent natural language processing capabilities. Nonetheless, these LLMs present many challenges, particularly in the …

EthicsFairness

Toward Reliable, Safe, and Secure LLMs for Scientific Applications

2026-03-18 · Saket Sanjeev Chaturvedi, Joshua Bergerson, Tanwi Mallick arxiv

As large language models (LLMs) evolve into autonomous "AI scientists," they promise transformative advances but introduce novel vulnerabilities, from potential "biosafety risks" to "dangerous explosions." Ensuring trust…

A Comprehensive Survey on the Trustworthiness of Large Language Models in Healthcare

2025-02-21 · Manar Aljohani, Jun Hou, Sindhura Kommu, Xuan Wang

The application of large language models (LLMs) in healthcare has the potential to revolutionize clinical decision-making, medical research, and patient care. As LLMs are increasingly integrated into healthcare systems, …

Decision MakingFairnessMisinformation

MolSafeEval: A Benchmark for Uncovering Safety Risks in AI-Generated Molecules

2026-07-01 · Tong Xu, Xinzhe Cao, Zhihui Zhu, Keyan Ding 외 arxiv

Current molecular generation benchmarks emphasize task complexity, molecule novelty, and property alignment; they largely overlook a critical concern: the potential safety risks of AI-generated molecules. In practice, ma…