Cross-Validated Loss-Based Covariance Matrix Estimator Selection in High Dimensions

Philippe Boileau,Nima S Hejazi,Mark J Van Der Laan,Sandrine Dudoit

doi:10.1080/10618600.2022.2110883

Abstract

The covariance matrix plays a fundamental role in many modern exploratory and inferential statistical procedures, including dimensionality reduction, hypothesis testing, and regression. In low-dimensional regimes, where the number of observations far exceeds the number of variables, the optimality of the sample covariance matrix as an estimator of this parameter is well-established. High-dimensional regimes do not admit such a convenience. Thus, a variety of estimators have been derived to overcome the shortcomings of the canonical estimator in such settings. Yet, selecting an optimal estimator from among the plethora available remains an open challenge. Using the framework of cross-validated loss-based estimation, we develop the theoretical underpinnings of just such an estimator selection procedure. We propose a general class of loss functions for covariance matrix estimation and establish accompanying finite-sample risk bounds and conditions for the asymptotic optimality of the cross-validation selector. In numerical experiments, we demonstrate the optimality of our proposed selector in moderate sample sizes and across diverse data-generating processes. The practical benefits of our procedure are highlighted in a dimension reduction application to single-cell transcriptome sequencing data.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Cross-Validated Loss-Based Covariance Matrix Estimator Selection in High Dimensions

Abstract

Talk to us

Similar Papers

More From: Journal of Computational and Graphical Statistics

Lead the way for us

Journal: Journal of Computational and Graphical Statistics	Publication Date: Aug 8, 2022
Citations: 2

Similar Papers

Ridge-type covariance and precision matrix estimators of the multivariate normal distribution
Wessel N Van Wieringen ... Gwenaël G R Leday
Statistical Papers | VOL. -
Wessel N Van Wieringen, et. al.Wessel N Van Wieringen ... Gwenaël G R Leday
21 Oct 2024
Statistical Papers | VOL. -

Eigenvalue-regularized covariance matrix estimators for high-dimensional data

-

01 Nov 2018
01 Nov 2018

Nonparametric estimation of large covariance matrices of longitudinal data
W B Wu
Biometrika | VOL. 90
W B WuW B Wu
01 Dec 2003
Biometrika | VOL. 90

Robust covariance and scatter matrix estimation under Huber’s contamination model
Mengjie Chen ... Chao Gao
The Annals of Statistics | VOL. 46
Mengjie Chen, et. al.Mengjie Chen ... Chao Gao
01 Oct 2018
The Annals of Statistics | VOL. 46

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Cross-Validated Loss-Based Covariance Matrix Estimator Selection in High Dimensions

Abstract

Talk to us

Similar Papers

More From: Journal of Computational and Graphical Statistics