paper-with-me

Papers

Chroma-VAE: Mitigating Shortcut Learning with Generative Classifiers

2022-11-28 · Wanqian Yang, Polina Kirichenko, Micah Goldblum, Andrew Gordon Wilson

Deep neural networks are susceptible to shortcut learning, using simple features to achieve low training loss without discovering essential semantic structure. Contrary to prior belief, we show that generative models alone are not sufficient to prevent shortcut learning, despite an incentive to recover a more comprehensive representation of the data than discriminative approaches. However, we observe that shortcuts are preferentially encoded with minimal information, a fact that generative models can exploit to mitigate shortcut learning. In particular, we propose Chroma-VAE, a two-pronged approach where a VAE classifier is initially trained to isolate the shortcut in a small latent subspace, allowing a secondary classifier to be trained on the complementary, shortcut-free latent subspace. In addition to demonstrating the efficacy of Chroma-VAE on benchmark and real-world shortcut learning tasks, our work highlights the potential for manipulating the latent space of generative classifiers to isolate or interpret specific correlations.

📄 PDF Abstract BibTeX arXiv:2211.15231

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Generative Classifiers Avoid Shortcut Solutions

2025-12-31 · Alexander C. Li, Ananya Kumar, Deepak Pathak arxiv

Discriminative approaches to classification often learn shortcuts that hold in-distribution but fail even under minor distribution shift. This failure mode stems from an overreliance on features that are spuriously corre…

KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models

2025-03-25 · Zhiwei Wang, Zhongxin Liu, Ying Li, Hongyu Sun 외

The emergence of large language models (LLMs) has significantly advanced the development of natural language processing (NLP), especially in text generation tasks like question answering. However, model hallucinations re…

HallucinationQuestion AnsweringText Generation

Learning to Detour: Shortcut Mitigating Augmentation for Weakly Supervised Semantic Segmentation

2024-05-28 · JuneHyoung Kwon, Eunju Lee, Yunsung Cho, Youngbin Kim

Weakly supervised semantic segmentation (WSSS) employing weak forms of labels has been actively studied to alleviate the annotation cost of acquiring pixel-level labels. However, classifiers trained on biased datasets te…

ObjectSemantic SegmentationWeakly supervised Semantic SegmentationWeakly-Supervised Semantic Segmentation

MeDi: Metadata-Guided Diffusion Models for Mitigating Biases in Tumor Classification

2025-06-20 · David Jacob Drexlin, Jonas Dippel, Julius Hense, Niklas Prenißl 외

Deep learning models have made significant advances in histological prediction tasks in recent years. However, for adaptation in clinical practice, their lack of robustness to varying conditions such as staining, scanner…

AMShortcut: An Inference- and Training-Efficient Inverse Design Model for Amorphous Materials

2026-03-31 · Yan Lin, Jonas A. Finkler, Tao Du, Jilin Hu 외 arxiv

Amorphous materials are solids that lack long-range atomic order but possess complex short- and medium-range order. Unlike crystalline materials that can be described by unit cells containing few up to hundreds of atoms,…