paper-with-me

홈 › Papers

Robust Learning via Golden Symmetric Loss of (un)Trusted Labels

2021-01-01 · Amirmasoud Ghiassi, Robert Birke, Lydia Y. Chen

Learning robust deep models against noisy labels becomes ever critical when today's data is commonly collected from open platforms and subject to adversarial corruption. The information on the label corruption process, i.e., corruption matrix, can greatly enhance the robustness of deep models but still fall behind in combating hard classes. In this paper, we propose to construct a golden symmetric loss (GSL) based on the estimated confusion matrix as to avoid overfitting to noisy labels and learn effectively from hard classes. GSL is the weighted sum of the corrected regular cross entropy and reverse cross entropy. By leveraging a small fraction of trusted clean data, we estimate the corruption matrix and use it to correct the loss as well as to determine the weights of GSL. We theoretically prove the robustness of the proposed loss function in the presence of dirty labels. We provide a heuristics to adaptively tune the loss weights of GSL according to the noise rate and diversity measured from the dataset. We evaluate our proposed golden symmetric loss on both vision and natural language deep models subject to different types of label noise patterns. Empirical results show that GSL can significantly outperform the existing robust training methods on different noise patterns, showing accuracy improvement up to 18% on CIFAR-100 and 1% on real world noisy dataset of Clothing1M.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TrustNet: Learning from Trusted Data Against (A)symmetric Label Noise

2020-07-13 · Amirmasoud Ghiassi, Taraneh Younesian, Robert Birke, Lydia Y. Chen

Robustness to label noise is a critical property for weakly-supervised classifiers trained on massive datasets. Robustness to label noise is a critical property for weakly-supervised classifiers trained on massive datase…

Golden Reference-Free Hardware Trojan Localization using Graph Convolutional Network

2022-07-14 · Rozhin Yasaei, Sina Faezi, Mohammad Abdullah Al Faruque

The globalization of the Integrated Circuit (IC) supply chain has moved most of the design, fabrication, and testing process from a single trusted entity to various untrusted third-party entities worldwide. The risk of u…

Feature Engineering

Trust-Aware Diversion for Data-Effective Distillation

2025-02-07 · Zhuojie Wu, Yanbin Liu, Xin Shen, Xiaofeng Cao 외

Dataset distillation compresses a large dataset into a small synthetic subset that retains essential information. Existing methods assume that all samples are perfectly labeled, limiting their real-world applications whe…

Dataset DistillationModel Optimization

Improving Event Temporal Relation Classification via Auxiliary Label-Aware Contrastive Learning

2022-10-01 · CCL 2022 10 · Sun Tiesen, Li Lishuang

“Event Temporal Relation Classification (ETRC) is crucial to natural language understanding. In recent years, the mainstream ETRC methods may not take advantage of lots of semantic information contained in golden tempora…

Contrastive LearningData AugmentationLanguage ModelingLanguage Modelling+4

Using Trusted Data to Train Deep Networks on Labels Corrupted by Severe Noise

2018-02-14 · NeurIPS 2018 12 · Dan Hendrycks, Mantas Mazeika, Duncan Wilson, Kevin Gimpel

The growing importance of massive datasets used for deep learning makes robustness to label noise a critical property for classifiers to have. Sources of label noise include automatic labeling, non-expert labeling, and l…

Data Poisoning