paper-with-me

홈 › Papers

It's easy to fool yourself: Case studies on identifying bias and confounding in bio-medical datasets

2019-12-12 · Subhashini Venugopalan, Arunachalam Narayanaswamy, Samuel Yang, Anton Geraschenko, Scott Lipnick, Nina Makhortova, James Hawrot, Christine Marques, Joao Pereira, Michael Brenner, Lee Rubin, Brian Wainger, Marc Berndl

Confounding variables are a well known source of nuisance in biomedical studies. They present an even greater challenge when we combine them with black-box machine learning techniques that operate on raw data. This work presents two case studies. In one, we discovered biases arising from systematic errors in the data generation process. In the other, we found a spurious source of signal unrelated to the prediction task at hand. In both cases, our prediction models performed well but under careful examination hidden confounders and biases were revealed. These are cautionary tales on the limits of using machine learning techniques on raw data from scientific experiments.

📄 PDF Abstract BibTeX arXiv:1912.07661

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningPrediction

Similar Papers 제목 키워드 기반

Is it a great Autonomous FX Trading Strategy or you are just fooling yourself

2021-01-15 · Murilo Sibrao Bernardini, Paulo Andre Lima de Castro

In this paper, we propose a method for evaluating autonomous trading strategies that provides realistic expectations, regarding the strategy's long-term performance. This method addresses This method addresses many pitfa…

Linguistic Cues of Deception in a Multilingual April Fools' Day Context

2021-11-06 · Katerina Papantoniou, Panagiotis Papadakos, Giorgos Flouris, Dimitris Plexousakis

In this work we consider the collection of deceptive April Fools' Day(AFD) news articles as a useful addition in existing datasets for deception detection tasks. Such collections have an established ground truth and are …

ArticlesDeception Detection

Tailoring Adversarial Attacks on Deep Neural Networks for Targeted Class Manipulation Using DeepFool Algorithm

2023-10-18 · S. M. Fazle Rabby Labib, Joyanta Jyoti Mondal, Meem Arafat Manab, Sarfaraz Newaz 외

The susceptibility of deep neural networks (DNNs) to adversarial attacks undermines their reliability across numerous applications, underscoring the necessity for an in-depth exploration of these vulnerabilities and the …

Seeing Through the Mask: Rethinking Adversarial Examples for CAPTCHAs

2024-09-09 · Yahya Jabary, Andreas Plesner, Turlan Kuzhagaliyev, Roger Wattenhofer

Modern CAPTCHAs rely heavily on vision tasks that are supposedly hard for computers but easy for humans. However, advances in image recognition models pose a significant threat to such CAPTCHAs. These models can easily b…

Don't Repeat Yourself: Stopping Verbatim Loops at Sampling Time

2026-08-24 · Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder, Sanjay Basu 외 arxiv

Large Language Models generate text autoregressively, but open-ended generation is prone to verbatim looping, in which models repeat spans already present in context. Standard defenses such as repetition, presence, and f…

Text Generation