paper-with-me

홈 › Papers

Testing for Underpowered Literatures

2024-06-19 · Stefan Faridani

How many experimental studies would have come to different conclusions had they been run on larger samples? I show how to estimate the expected number of statistically significant results that a set of experiments would have reported had their sample sizes all been counterfactually increased. The proposed deconvolution estimator is asymptotically normal and adjusts for publication bias. Unlike related methods, this approach requires no assumptions of any kind about the distribution of true intervention treatment effects. An application to randomized trials (RCTs) published in economics journals finds that doubling every sample would increase the power of t-tests by 7.2 percentage points on average. This effect is smaller than for non-RCTs and comparable to systematic replications in laboratory psychology where previous studies enabled more accurate power calculations. This suggests that RCTs are on average relatively insensitive to sample size increases. Funders should generally consider sponsoring more experiments rather than fewer, larger ones.

📄 PDF Abstract BibTeX arXiv:2406.13122

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

A Multi-Evidence Framework Rescues Low-Power Prognostic Signals and Rejects Statistical Artifacts in Cancer Genomics

2025-10-21 · Gokturk Aytug Akarlar arxiv

Motivation: Standard genome-wide association studies in cancer genomics rely on statistical significance with multiple testing correction, but systematically fail in underpowered cohorts. In TCGA breast cancer (n=967, 13…

Causal Inference

With Little Power Comes Great Responsibility

2020-10-13 · EMNLP 2020 11 · Dallas Card, Peter Henderson, Urvashi Khandelwal, Robin Jia 외

Despite its importance to experimental design, statistical power (the probability that, given a real effect, an experiment will reject the null hypothesis) has largely been ignored by the NLP community. Underpowered expe…

Experimental DesignMachine TranslationTranslation

A reproducible effect size is more useful than an irreproducible hypothesis test to analyze high throughput sequencing datasets

2019-05-13

Motivation: P values derived from the null hypothesis significance testing framework are strongly affected by sample size, and are known to be irreproducible in underpowered studies, yet no suitable replacement has been …

TAG

AdaPT-GMM: Powerful and robust covariate-assisted multiple testing

2021-06-30 · Patrick Chao, William Fithian

We propose a new empirical Bayes method for covariate-assisted multiple testing with false discovery rate (FDR) control, where we model the local false discovery rate for each hypothesis as a function of both its covaria…

Neural Metaphor Detecting with CNN-LSTM Model

2018-06-01 · WS 2018 6 · Chuhan Wu, Fangzhao Wu, Yubo Chen, Sixing Wu 외

Metaphors are figurative languages widely used in daily life and literatures. It{'}s an important task to detect the metaphors evoked by texts. Thus, the metaphor shared task is aimed to extract metaphors from plain text…

Machine TranslationmodelPOSSentiment Analysis