arrow
Return

Clustering, Spatial Correlations, and Randomization Inference

delete2012-06-01
delete78
delete
OA
AI
T
Thomas Barrios *
R
Rebecca Diamond
G
Guido W. Imbens
M
Michal Kolesár
DOI:10.1080/01621459.2012.682524delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
It is a standard practice in regression analyses to allow for clustering in the error covariance matrix if the explanatory variable of interest varies at a more aggregate level (e.g., the state level) than the units of observation (e.g., individuals). Often, however, the structure of the error covariance matrix is more complex, with correlations not vanishing for units in different clusters. Here, we explore the implications of such correlations for the actual and estimated precision of least squares estimators. Our main theoretical result is that with equal-sized clusters, if the covariate of interest is randomly assigned at the cluster level, only accounting for nonzero covariances at the cluster level, and ignoring correlations between clusters as well as differences in within-cluster correlations, leads to valid confidence intervals. However, in the absence of random assignment of the covariates, ignoring general correlation structures may lead to biases in standard errors. We illustrate our findings using the 5% public-use census data. Based on these results, we recommend that researchers, as a matter of routine, explore the extent of spatial correlations in explanatory variables beyond state-level clustering.
Keywords:
Clustered standard errors
Confidence intervals
Misspecification
Random assignment

Journal

J
Journal of the American Statistical Association
IF:
3
Papers:
5.1K
Citations:
4.8W

Organization

H
Harvard University
Scholars:
26.5W
Papers: 22.0W
Citations: 28.7W