paper-with-me

홈 › Papers

When does Privileged Information Explain Away Label Noise?

2023-03-03 · Guillermo Ortiz-Jimenez, Mark Collier, Anant Nawalgaria, Alexander D'Amour, Jesse Berent, Rodolphe Jenatton, Effrosyni Kokiopoulou

Leveraging privileged information (PI), or features available during training but not at test time, has recently been shown to be an effective method for addressing label noise. However, the reasons for its effectiveness are not well understood. In this study, we investigate the role played by different properties of the PI in explaining away label noise. Through experiments on multiple datasets with real PI (CIFAR-N/H) and a new large-scale benchmark ImageNet-PI, we find that PI is most helpful when it allows networks to easily distinguish clean from noisy data, while enabling a learning shortcut to memorize the noisy examples. Interestingly, when PI becomes too predictive of the target label, PI methods often perform worse than their no-PI baselines. Based on these findings, we propose several enhancements to the state-of-the-art PI methods and demonstrate the potential of PI as a means of tackling label noise. Finally, we show how we can easily combine the resulting PI approaches with existing no-PI techniques designed to deal with label noise.

📄 PDF Abstract BibTeX arXiv:2303.01806

Code (1)

google-research-datasets/imagenet_pi 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Transfer and Marginalize: Explaining Away Label Noise with Privileged Information

2022-02-18 · Mark Collier, Rodolphe Jenatton, Efi Kokiopoulou, Jesse Berent

Supervised learning datasets often have privileged information, in the form of features which are available at training time but are not available at test time e.g. the ID of the annotator that provided the label. We arg…

TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment

2026-05-11 · Jiaxuan Wang, Xuan Ouyang, Zhiyu Chen, Yulan Hu 외 arxiv

On-policy self-distillation (self-OPD) densifies reinforcement learning with verifiable rewards (RLVR) by letting a policy teach itself under privileged context. We find that when this guidance spans the full response, a…

Reinforcement Learning

Attention that does not Explain Away

2020-09-29 · Nan Ding, Xinjie Fan, Zhenzhong Lan, Dale Schuurmans 외

Models based on the Transformer architecture have achieved better accuracy than the ones based on competing architectures for a large set of tasks. A unique feature of the Transformer is its universal application of a se…

The Variational Bandwidth Bottleneck: Stochastic Evaluation on an Information Budget

2020-04-24 · ICLR 2020 1 · Anirudh Goyal, Yoshua Bengio, Matthew Botvinick, Sergey Levine

In many applications, it is desirable to extract only the relevant information from complex input data, which involves making a decision about which input features are relevant. The information bottleneck method formaliz…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Variational Inference

A Tutorial on Principal Component Analysis

2014-04-03 · Jonathon Shlens

Principal component analysis (PCA) is a mainstay of modern data analysis - a black box that is widely used but (sometimes) poorly understood. The goal of this paper is to dispel the magic behind this black box. This manu…