paper-with-me

Papers

Detecting Dataset Bias in Medical AI: A Generalized and Modality-Agnostic Auditing Framework

2025-03-13 · Nathan Drenkow, Mitchell Pavlak, Keith Harrigian, Ayah Zirikly, Adarsh Subbaswamy, Mathias Unberath

Data-driven AI is establishing itself at the center of evidence-based medicine. However, reports of shortcomings and unexpected behavior are growing due to AI's reliance on association-based learning. A major reason for this behavior: latent bias in machine learning datasets can be amplified during training and/or hidden during testing. We present a data modality-agnostic auditing framework for generating targeted hypotheses about sources of bias which we refer to as Generalized Attribute Utility and Detectability-Induced bias Testing (G-AUDIT) for datasets. Our method examines the relationship between task-level annotations and data properties including protected attributes (e.g., race, age, sex) and environment and acquisition characteristics (e.g., clinical site, imaging protocols). G-AUDIT automatically quantifies the extent to which the observed data attributes may enable shortcut learning, or in the case of testing data, hide predictions made based on spurious associations. We demonstrate the broad applicability and value of our method by analyzing large-scale medical datasets for three distinct modalities and learning tasks: skin lesion classification in images, stigmatizing language classification in Electronic Health Records (EHR), and mortality prediction for ICU tabular data. In each setting, G-AUDIT successfully identifies subtle biases commonly overlooked by traditional qualitative methods that focus primarily on social and ethical objectives, underscoring its practical value in exposing dataset-level risks and supporting the downstream development of reliable AI systems. Our method paves the way for achieving deeper understanding of machine learning datasets throughout the AI development life-cycle from initial prototyping all the way to regulation, and creates opportunities to reduce model bias, enabling safer and more trustworthy AI systems.

📄 PDF Abstract BibTeX arXiv:2503.09969

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeLesion ClassificationMortality PredictionSkin Lesion Classification

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

A Causal Approach to Mitigate Modality Preference Bias in Medical Visual Question Answering

2025-05-22 · Shuchang Ye, Usman Naseem, Mingyuan Meng, Dagan Feng 외

Medical Visual Question Answering (MedVQA) is crucial for enhancing the efficiency of clinical diagnosis by providing accurate and timely responses to clinicians' inquiries regarding medical images. Existing MedVQA model…

counterfactualMedical Visual Question AnsweringQuestion AnsweringVisual Question Answering+1

Dual-level Modality Debiasing Learning for Unsupervised Visible-Infrared Person Re-Identification

2025-12-03 · Jiaze Li, Yan Lu, Bin Liu, Guojun Yin 외 arxiv

Two-stage learning pipeline has achieved promising results in unsupervised visible-infrared person re-identification (USL-VI-ReID). It first performs single-modality learning and then operates cross-modality learning to …

Person Re-Identification

On the Risk of Misleading Reports: Diagnosing Textual Biases in Multimodal Clinical AI

2025-07-31 · David Restrepo, Ira Ktena, Maria Vakalopoulou, Stergios Christodoulidis 외 arxiv

Clinical decision-making relies on the integrated analysis of medical images and the associated clinical reports. While Vision-Language Models (VLMs) can offer a unified framework for such tasks, they can exhibit strong …

Binary Classification

ADAPT: Multimodal Learning for Detecting Physiological Changes under Missing Modalities

2024-07-04 · Julie Mordacq, Leo Milecki, Maria Vakalopoulou, Steve Oudot 외

Multimodality has recently gained attention in the medical domain, where imaging or video modalities may be integrated with biomedical signals or health records. Yet, two challenges remain: balancing the contributions of…

Cross-Modality Fourier Feature for Medical Image Synthesis

2023-07-10 · journal 2023 7 · Mei Ma; Ling Lin; Heng Wang; Zhendong Li; Hao Liu

In this paper, we propose a cross-modality fourier feature (CMFF) method via frequency selection, which learns the rational anatomical structure for targeting medical modality images. Unlike existing works seeking pixel-…

Image Generation