vix.ing · top · new · best · stats · spec

Reclustering: A New Method to Test the Appropriate Level of Clustering

2025/11/11 by Kentaro Fukumoto, Fukumoto, Kentaro
Mathematics · #Advanced Statistical Methods and Models #Census and Population Estimation #FOS: Computer and information sciences #Methodology (stat.ME) #Statistical Methods and Bayesian Inference

paper · pdf · doi:10.48550/arxiv.2511.08184

openalex publication_date 2025/11/11 · openalex created_date 2025/11/13 · openalex updated_date 2026/07/28

Abstract

When scholars suspect units are dependent on each other within clusters but independent of each other across clusters, they employ cluster-robust standard errors (CRSEs). Nevertheless, what to cluster over is sometimes unknown. For instance, in the case of cross-sectional survey samples, clusters may be households, municipalities, counties, or states. A few approaches have been proposed, although they are based on asymptotics. I propose a new method to address this issue that works in a finite sample: reclustering. That is, we randomly and repeatedly group fine clusters into new gross clusters and calculate a statistic such as CRSEs. Under the null hypothesis that fine clusters are independent of each other, how they are grouped into gross clusters should not matter for any cluster-sensitive statistic. Thus, if the statistic based on the original clustering is a significant outlier against the distributions of the statistics induced by reclustering, it is reasonable to reject the null hypothesis and employ gross clusters. I compare the performance of reclustering with that of a few previous tests using Monte Carlo simulation and application.

Citations

Related