paper-with-me

홈 › Papers

Quantifying Spuriousness of Biased Datasets Using Partial Information Decomposition

2024-06-29 · Barproda Halder, Faisal Hamman, Pasan Dissanayake, Qiuyi Zhang, Ilia Sucholutsky, Sanghamitra Dutta

Spurious patterns refer to a mathematical association between two or more variables in a dataset that are not causally related. However, this notion of spuriousness, which is usually introduced due to sampling biases in the dataset, has classically lacked a formal definition. To address this gap, this work presents the first information-theoretic formalization of spuriousness in a dataset (given a split of spurious and core features) using a mathematical framework called Partial Information Decomposition (PID). Specifically, we disentangle the joint information content that the spurious and core features share about another target variable (e.g., the prediction label) into distinct components, namely unique, redundant, and synergistic information. We propose the use of unique information, with roots in Blackwell Sufficiency, as a novel metric to formally quantify dataset spuriousness and derive its desirable properties. We empirically demonstrate how higher unique information in the spurious features in a dataset could lead a model into choosing the spurious features over the core features for inference, often having low worst-group-accuracy. We also propose a novel autoencoder-based estimator for computing unique information that is able to handle high-dimensional image data. Finally, we also show how this unique information in the spurious feature is reduced across several dataset-based spurious-pattern-mitigation techniques such as data reweighting and varying levels of background mixing, demonstrating a novel tradeoff between unique information (spuriousness) and worst-group-accuracy.

📄 PDF Abstract BibTeX arXiv:2407.00482

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Spuriousness-Aware Meta-Learning for Learning Robust Classifiers

2024-06-15 · Guangtao Zheng, Wenqian Ye, Aidong Zhang

Spurious correlations are brittle associations between certain attributes of inputs and target variables, such as the correlation between an image background and an object class. Deep image classifiers often leverage the…

AttributeLanguage ModellingMeta-Learning

Quantifying and Learning Static vs. Dynamic Information in Deep Spatiotemporal Networks

2022-11-03 · Matthew Kowal, Mennatullah Siam, Md Amirul Islam, Neil D. B. Bruce 외

There is limited understanding of the information captured by deep spatiotemporal models in their intermediate representations. For example, while evidence suggests that action recognition algorithms are heavily influenc…

Action RecognitionInstance SegmentationSemantic SegmentationVideo Instance Segmentation+2

The Multiple Dimensions of Spuriousness in Machine Learning

2024-11-07 · Samuel J. Bell, Skyler Wang

Learning correlations from data forms the foundation of today's machine learning (ML) and artificial intelligence (AI) research. While such an approach enables the automatic discovery of patterned relationships within bi…

Fairness

A Deeper Dive Into What Deep Spatiotemporal Networks Encode: Quantifying Static vs. Dynamic Information

2022-06-06 · CVPR 2022 1 · Matthew Kowal, Mennatullah Siam, Md Amirul Islam, Neil D. B. Bruce 외

Deep spatiotemporal models are used in a variety of computer vision tasks, such as action recognition and video object segmentation. Currently, there is a limited understanding of what information is captured by these mo…

Action RecognitionSemantic SegmentationVideo Object SegmentationVideo Semantic Segmentation

Towards Trustworthy Explanation: On Causal Rationalization

2023-06-25 · Wenbo Zhang, Tong Wu, Yunlong Wang, Yong Cai 외

With recent advances in natural language processing, rationalization becomes an essential self-explaining diagram to disentangle the black box by selecting a subset of input texts to account for the major variation in pr…

Causal Inference