paper-with-me

Papers

The Clever Hans Effect in Unsupervised Learning

2024-08-15 · Jacob Kauffmann, Jonas Dippel, Lukas Ruff, Wojciech Samek, Klaus-Robert Müller, Grégoire Montavon

Unsupervised learning has become an essential building block of AI systems. The representations it produces, e.g. in foundation models, are critical to a wide variety of downstream applications. It is therefore important to carefully examine unsupervised models to ensure not only that they produce accurate predictions, but also that these predictions are not "right for the wrong reasons", the so-called Clever Hans (CH) effect. Using specially developed Explainable AI techniques, we show for the first time that CH effects are widespread in unsupervised learning. Our empirical findings are enriched by theoretical insights, which interestingly point to inductive biases in the unsupervised learning machine as a primary source of CH effects. Overall, our work sheds light on unexplored risks associated with practical applications of unsupervised learning and suggests ways to make unsupervised learning more robust.

📄 PDF Abstract BibTeX arXiv:2408.08041

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Clever Hans Effect in Anomaly Detection

2020-06-18 · Jacob Kauffmann, Lukas Ruff, Grégoire Montavon, Klaus-Robert Müller

The 'Clever Hans' effect occurs when the learned model produces correct predictions based on the 'wrong' features. This effect which undermines the generalization capability of an ML model and goes undetected by standard…

Anomaly DetectionExplainable Artificial Intelligence (XAI)Outlier Detection

Preemptively Pruning Clever-Hans Strategies in Deep Neural Networks

2023-04-12 · Lorenz Linhardt, Klaus-Robert Müller, Grégoire Montavon

Robustness has become an important consideration in deep learning. With the help of explainable AI, mismatches between an explained model's decision strategy and the user's domain knowledge (e.g. Clever Hans effects) hav…

Clever Hans Effect Found in Automatic Detection of Alzheimer's Disease through Speech

2024-06-11 · Yin-Long Liu, Rui Feng, Jia-Hong Yuan, Zhen-Hua Ling

We uncover an underlying bias present in the audio recordings produced from the picture description task of the Pitt corpus, the largest publicly accessible database for Alzheimer's Disease (AD) detection research. Even …

Technical Report on the CleverHans v2.1.0 Adversarial Examples Library

2016-10-03 · Nicolas Papernot, Fartash Faghri, Nicholas Carlini, Ian Goodfellow 외

An adversarial example library for constructing attacks, building defenses, and benchmarking both

Adversarial AttackAdversarial DefenseBenchmarking

Making deep neural networks right for the right scientific reasons by interacting with their explanations

2020-01-15 · Patrick Schramowski, Wolfgang Stammer, Stefano Teso, Anna Brugger 외

Deep neural networks have shown excellent performances in many real-world applications. Unfortunately, they may show "Clever Hans"-like behavior -- making use of confounding factors within datasets -- to achieve high per…

BIG-bench Machine LearningPlant Phenotyping