arXiv · 2205.07950
The Power of Tests for Detecting $p$-Hacking
Abstract
A flourishing empirical literature investigates the prevalence of $p$-hacking based on the distribution of $p$-values across studies. Interpreting results in this literature requires a careful understanding of the power of methods for detecting $p$-hacking. We theoretically study the implications of likely forms of $p$-hacking on the distribution of $p$-values to understand the power of tests for detecting it. Power can be low and depends crucially on the $p$-hacking strategy and the distribution of true effects. Combined tests for upper bounds and monotonicity and tests for continuity of the $p$-curve tend to have the highest power for detecting $p$-hacking.
Explore related subjects
Keep this discovery
Graham Elliott, Nikolay Kudrin, Kaspar Wüthrich. 2022-05-16. The Power of Tests for Detecting $p$-Hacking. https://arxiv.org/abs/2205.07950
Cite the original work for its findings. Save a collection to share your selection of sources.