2016/10/31 by Taras Bodnar, Ostap Okhrin, Nestor Parolya · 2 citations
Economics, Econometrics and Finance · Mathematics · #math.ST #q-fin.ST #stat.TH
paper · pdf · doi:10.1016/j.jmva.2018.07.004
20 pages, UPDATE2: revised version of the manuscript accepted for publication in Journal of Multivariate Analysis
arxiv created 2018/07/14 · arxiv updated 2018/07/17
In this paper we derive the optimal linear shrinkage estimator for the high-dimensional mean vector using random matrix theory. The results are obtained under the assumption that both the dimension p and the sample size n tend to infinity in such a way that p/n → c∈(0,∞). Under weak conditions imposed on the underlying data generating mechanism, we find the asymptotic equivalents to the optimal shrinkage intensities and estimate them consistently. The proposed nonparametric estimator for the high-dimensional mean vector has a simple structure and is proven to minimize asymptotically, with probability 1, the quadratic loss when c∈(0,1). When c∈(1, ∞) we modify the estimator by using a feasible estimator for the precision covariance matrix. To this end, an exhaustive simulation study and an application to real data are provided where the proposed estimator is compared with known benchmarks from the literature. It turns out that the existing estimators of the mean vector, including the new proposal, converge to the sample mean vector when the true mean vector has an unbounded Euclidean norm.