paper-with-me

홈 › Papers

Observational Multiplicity

2025-07-30 · Erin George, Deanna Needell, Berk Ustun arxiv

Many prediction tasks can admit multiple models that can perform almost equally well. This phenomenon can can undermine interpretability and safety when competing models assign conflicting predictions to individuals. In this work, we study how arbitrariness can arise in probabilistic classification tasks as a result of an effect that we call \emph{observational multiplicity}. We discuss how this effect arises in a broad class of practical applications where we learn a classifier to predict probabilities $p_i \in [0,1]$ but are given a dataset of observations $y_i \in \{0,1\}$. We propose to evaluate the arbitrariness of individual probability predictions through the lens of \emph{regret}. We introduce a measure of regret for probabilistic classification tasks, which measures how the predictions of a model could change as a result of different training labels change. We present a general-purpose method to estimate the regret in a probabilistic classification task. We use our measure to show that regret is higher for certain groups in the dataset and discuss potential applications of regret. We demonstrate how estimating regret promote safety in real-world applications by abstention and data collection.

📄 PDF Abstract BibTeX arXiv:2507.23136

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Decomposing Observational Multiplicity in Decision Trees: Leaf and Structural Regret

2026-03-12 · Mustafa Cavus arxiv

Many machine learning tasks admit multiple models that perform almost equally well, a phenomenon known as predictive multiplicity. A fundamental source of this multiplicity is observational multiplicity, which arises fro…

Data as a Lever: A Neighbouring Datasets Perspective on Predictive Multiplicity

2025-10-24 · Prakhar Ganesh, Hsiang Hsu, Golnoosh Farnadi arxiv

Multiplicity, the existence of equally good yet competing models, has received growing attention in recent years. While prior work has emphasized modelling choices, the critical role of data in shaping multiplicity has b…

Active Learning

An Empirical Investigation into Benchmarking Model Multiplicity for Trustworthy Machine Learning: A Case Study on Image Classification

2023-11-24 · Prakhar Ganesh

Deep learning models have proven to be highly successful. Yet, their over-parameterization gives rise to model multiplicity, a phenomenon in which multiple models achieve similar performance but exhibit distinct underlyi…

Benchmarkingimage-classificationImage ClassificationModel Selection

Multiplicity is an Inevitable and Inherent Challenge in Multimodal Learning

2025-05-26 · Sanghyuk Chun

Multimodal learning has seen remarkable progress, particularly with the emergence of large-scale pre-training across various modalities. However, most current approaches are built on the assumption of a deterministic, on…

Position

Mitigating the Multiplicity Burden: The Role of Calibration in Reducing Predictive Multiplicity of Classifiers

2026-03-12 · Mustafa Cavus arxiv

As machine learning models are increasingly deployed in high-stakes environments, ensuring both probabilistic reliability and prediction stability has become critical. This paper examines the interplay between classifica…