paper-with-me

홈 › Papers

Risk-Calibrated Learning: Minimizing Fatal Errors in Medical AI

2026-04-14 · Abolfazl Mohammadi-Seif, Ricardo Baeza-Yates arxiv

Deep learning models often achieve expert-level accuracy in medical image classification but suffer from a critical flaw: semantic incoherence. These high-confidence mistakes that are semantically incoherent (e.g., classifying a malignant tumor as benign) fundamentally differ from acceptable errors which stem from visual ambiguity. Unlike safe, fine-grained disagreements, these fatal failures erode clinical trust. To address this, we propose Risk-Calibrated Learning, a technique that explicitly distinguishes between visual ambiguity (fine-grained errors) and catastrophic structural errors. By embedding a confusion-aware clinical severity matrix M into the optimization landscape, our method suppresses critical errors (false negatives) without requiring complex architectural changes. We validate our approach in four different imaging modalities: Brain Tumor MRI, ISIC 2018 (Dermoscopy), BreaKHis (Breast Histopathology), and SICAPv2 (Prostate Histopathology). Extensive experiments demonstrate that our Risk-Calibrated Loss consistently reduces the Critical Error Rate (CER) for all four datasets, achieving relative safety improvements ranging from 20.0% (on breast histopathology) to 92.4% (on prostate histopathology) compared to state-of-the-art baselines such as Focal Loss. These results confirm that our method offers a superior safety-accuracy trade-off across both CNN and Transformer architectures.

📄 PDF Abstract BibTeX arXiv:2604.12693

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Image Classification

Similar Papers 제목 키워드 기반

Do not forget interaction: Predicting fatality of COVID-19 patients using logistic regression

2020-06-30 · Feng Zhou, Tao Chen, Baiying Lei

Amid the ongoing COVID-19 pandemic, whether COVID-19 patients with high risks can be recovered or not depends, to a large extent, on how early they will be treated appropriately before irreversible consequences are cause…

regression

CARE: A Conformal Safety Layer for Medical Summarization

2026-06-08 · Suhana Bedi, Bridget Lin, Anson Y. Zhou, Chloe O. Stanwyck 외 arxiv

Large language models (LLMs) are increasingly used for medical summarization, but their outputs can omit medically important information and introduce unsupported claims. Existing error-detection methods produce heuristi…

Can LLMs Accurately Score Medical Diagnoses and Clinical Reasoning?

2026-04-16 · Amy Rouillard, Sitwala Mundia, Linda Camara, Ziyaad Dangor 외 arxiv

Evaluating medical AI systems using expert clinician panels is costly and slow, motivating the use of large language models (LLMs) as alternative adjudicators. Here, we evaluate an LLM Jury, composed of three frontier AI…

DOMINO: Domain-aware Loss for Deep Learning Calibration

2023-02-10 · Skylar E. Stolte, Kyle Volle, Aprinda Indahlastari, Alejandro Albizu 외

Deep learning has achieved the state-of-the-art performance across medical imaging tasks; however, model calibration is often not considered. Uncalibrated models are potentially dangerous in high-risk applications since …

Deep Learning

Crash Report Data Analysis for Creating Scenario-Wise, Spatio-Temporal Attention Guidance to Support Computer Vision-based Perception of Fatal Crash Risks

2021-09-06 · Yu Li, Muhammad Monjurul Karim, Ruwen Qin

Reducing traffic fatalities and serious injuries is a top priority of the US Department of Transportation. The computer vision (CV)-based crash anticipation in the near-crash phase is receiving growing attention. The abi…

Clustering