paper-with-me

홈 › Papers

Are you doing better than random guessing? A call for using negative controls when evaluating causal discovery algorithms

2024-12-13 · Anne Helby Petersen

New proposals for causal discovery algorithms are typically evaluated using simulations and a few selected real data examples with known data generating mechanisms. However, there does not exist a general guideline for how such evaluation studies should be designed, and therefore, comparing results across different studies can be difficult. In this article, we propose to use negative controls as a common evaluation baseline by posing the question: Are we doing better than random guessing? For the task of graph skeleton estimation, we derive exact distributional results under random guessing for the expected behavior of a range of typical causal discovery evaluation metrics, including precision and recall. We show that these metrics can achieve very favorable values under random guessing in certain scenarios, and hence warn against using them without also reporting negative control results, i.e., performance under random guessing. We also propose an exact test of overall skeleton fit, and showcase its use on a real data application. Finally, we propose a general pipeline for using negative controls beyond the skeleton estimation task, and apply it both in a simulated example and a real data application.

📄 PDF Abstract BibTeX arXiv:2412.10039

Code (0)

등록된 구현이 없습니다.

Tasks

Causal Discovery

Similar Papers 제목 키워드 기반

Provable Weak-to-Strong Generalization via Benign Overfitting

2024-10-06 · David X. Wu, Anant Sahai

The classic teacher-student model in machine learning posits that a strong teacher supervises a weak student to improve the student's capabilities. We instead consider the inverted situation, where a weak teacher supervi…

Inferring Convolutional Neural Networks' accuracies from their architectural characterizations

2020-01-07 · Duc Hoang, Jesse Hamer, Gabriel N. Perdue, Steven R. Young 외

Convolutional Neural Networks (CNNs) have shown strong promise for analyzing scientific data from many domains including particle imaging detectors. However, the challenge of choosing the appropriate network architecture…

Model Selection

The perils of being unhinged: On the accuracy of classifiers minimizing a noise-robust convex loss

2021-12-08 · Philip M. Long, Rocco A. Servedio

Van Rooyen et al. introduced a notion of convex loss functions being robust to random classification noise, and established that the "unhinged" loss function is robust in this sense. In this note we study the accuracy of…

Arbitrarily Large Labelled Random Satisfiability Formulas for Machine Learning Training

2022-11-21 · Dimitris Achlioptas, Amrit Daswaney, Periklis A. Papakonstantinou

Applying deep learning to solve real-life instances of hard combinatorial problems has tremendous potential. Research in this direction has focused on the Boolean satisfiability (SAT) problem, both because of its theoret…

Combinatorial Optimization

Mapping Between fMRI Responses to Movies and their Natural Language Annotations

2016-10-13 · Kiran Vodrahalli, Po-Hsuan Chen, YIngyu Liang, Christopher Baldassano 외

Several research groups have shown how to correlate fMRI responses to the meanings of presented stimuli. This paper presents new methods for doing so when only a natural language annotation is available as the descriptio…

Scene ClassificationSentenceSentence EmbeddingSentence-Embedding