paper-with-me

홈 › Papers

On Arbitrary Predictions from Equally Valid Models

2025-07-25 · Sarah Lockfisch, Kristian Schwethelm, Martin Menten, Rickmer Braren, Daniel Rueckert, Alexander Ziller, Georgios Kaissis arxiv

Model multiplicity refers to the existence of multiple machine learning models that describe the data equally well but may produce different predictions on individual samples. In medicine, these models can admit conflicting predictions for the same patient -- a risk that is poorly understood and insufficiently addressed. In this study, we empirically analyze the extent, drivers, and ramifications of predictive multiplicity across diverse medical tasks and model architectures, and show that even small ensembles can mitigate/eliminate predictive multiplicity in practice. Our analysis reveals that (1) standard validation metrics fail to identify a uniquely optimal model and (2) a substantial amount of predictions hinges on arbitrary choices made during model development. Using multiple models instead of a single model reveals instances where predictions differ across equally plausible models -- highlighting patients that would receive arbitrary diagnoses if any single model were used. In contrast, (3) a small ensemble paired with an abstention strategy can effectively mitigate measurable predictive multiplicity in practice; predictions with high inter-model consensus may thus be amenable to automated classification. While accuracy is not a principled antidote to predictive multiplicity, we find that (4) higher accuracy achieved through increased model capacity reduces predictive multiplicity. Our findings underscore the clinical importance of accounting for model multiplicity and advocate for ensemble-based strategies to improve diagnostic reliability. In cases where models fail to reach sufficient consensus, we recommend deferring decisions to expert review.

📄 PDF Abstract BibTeX arXiv:2507.19408

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Online Multivalid Learning: Means, Moments, and Prediction Intervals

2021-01-05 · Varun Gupta, Christopher Jung, Georgy Noarov, Mallesh M. Pai 외

We present a general, efficient technique for providing contextual predictions that are "multivalid" in various senses, against an online sequence of adversarially chosen examples $(x,y)$. This means that the resulting e…

Conformal PredictionPredictionPrediction Intervals

A Unified Model for Spatio-Temporal Prediction Queries with Arbitrary Modifiable Areal Units

2024-03-10 · Liyue Chen, Jiangyi Fang, Tengfei Liu, Shaosheng Cao 외

Spatio-Temporal (ST) prediction is crucial for making informed decisions in urban location-based applications like ride-sharing. However, existing ST models often require region partition as a prerequisite, resulting in …

Prediction

Aggregating Predictions on Multiple Non-disclosed Datasets using Conformal Prediction

2018-06-11 · Ola Spjuth, Lars Carlsson, Niharika Gauraha

Conformal Prediction is a machine learning methodology that produces valid prediction regions under mild conditions. In this paper, we explore the application of making predictions over multiple data sources of different…

Conformal PredictionPredictionvalid

Using Weighted P-Values in Fisher's Method

2020-06-17 · Arvind Thiagarajan

Fisher's method prescribes a way to combine p-values from multiple experiments into a single p-value. However, the original method can only determine a combined p-value analytically if all constituent p-values are weight…

A statistical framework for fair predictive algorithms

2016-10-25 · Kristian Lum, James Johndrow

Predictive modeling is increasingly being employed to assist human decision-makers. One purported advantage of replacing human judgment with computer models in high stakes settings-- such as sentencing, hiring, policing,…