Feature Noise Induces Loss Discrepancy Across Groups
The performance of standard learning procedures has been observed to differ widely across groups. Recent studies usually attribute this loss discrepancy to an information deficiency for one group (e.g., one group has less data). In this work, we point to a more subtle source of loss discrepancy---feature noise. Our main result is that even when there is no information deficiency specific to one group (e.g., both groups have infinite data), adding the same amount of feature noise to all individuals leads to loss discrepancy. For linear regression, we thoroughly characterize the effect of feature noise on loss discrepancy in terms of the amount of noise, the difference between moments of the two groups, and whether group information is used or not. We then show this loss discrepancy does not vanish immediately if a shift in distribution causes the groups to have similar moments. On three real-world datasets, we show feature noise increases the loss discrepancy if groups have different distributions, while it does not affect the loss discrepancy on datasets where groups have similar distributions.
Code (1)
Tasks
AttributeSimilar Papers 제목 키워드 기반
Dual Domain-Adversarial Learning for Audio-Visual Saliency Prediction
Both visual and auditory information are valuable to determine the salient regions in videos. Deep convolution neural networks (CNN) showcase strong capacity in coping with the audio-visual saliency prediction task. Due …
Domain AdaptationPredictionSaliency PredictionUnsupervised Domain AdaptationA Novel Bi-hemispheric Discrepancy Model for EEG Emotion Recognition
The neuroscience study has revealed the discrepancy of emotion expression between left and right hemispheres of human brain. Inspired by this study, in this paper, we propose a novel bi-hemispheric discrepancy model (BiH…
EEGEEG Emotion RecognitionElectroencephalogram (EEG)Emotion RecognitionCalibrating Teacher--Student Discrepancy for On-Policy Distillation
On-policy distillation (OPD) improves reasoning models by learning the token-level discrepancy between a stronger teacher and an on-policy student. However, this discrepancy does not purely reflect the capability gap bet…
Mathematical ReasoningRethinking Benign Overfitting in Two-Layer Neural Networks
Recent theoretical studies (Kou et al., 2023; Cao et al., 2022) have revealed a sharp phase transition from benign to harmful overfitting when the noise-to-feature ratio exceeds a threshold-a situation common in long-tai…
MemorizationAdversarial Discriminative Heterogeneous Face Recognition
The gap between sensing patterns of different face modalities remains a challenging problem in heterogeneous face recognition (HFR). This paper proposes an adversarial discriminative feature learning framework to close t…
Face HallucinationFace RecognitionHallucinationHeterogeneous Face Recognition