vix.ing · top · new · best · stats

Channeling Fisher: Randomization Tests and the Statistical Insignificance of Seemingly Significant Experimental Results*

2018/11/19 by Alwyn Young · 616 citations
Mathematics · Psychology · #Biology #Clinical trial #Econometrics #Insignificance #Mathematics #Psychology #Randomization #Social psychology #Statistical Distribution Estimation and Applications #Statistical Methods and Bayesian Inference #Statistical Methods and Inference #Statistical analysis #Statistical hypothesis testing #Statistics

paper · open access · doi:10.1093/qje/qjy029

published in The Quarterly Journal of Economics 134(2), 557-598 (Oxford University Press)

openalex publication_date 2018/11/19 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/30

Abstract

I follow R. A. Fisher's,The Design of Experiments (1935), using randomization statistical inference to test the null hypothesis of no treatment effects in a comprehensive sample of 53 experimental papers drawn from the journals of the American Economic Association. In the average paper, randomization tests of the significance of individual treatment effects find 13% to 22% fewer significant results than are found using authors’ methods. In joint tests of multiple treatment effects appearing together in tables, randomization tests yield 33% to 49% fewer statistically significant results than conventional tests. Bootstrap and jackknife methods support and confirm the randomization results.

Citations

Cited by

Related