SOTAVerified

Minimax sparse principal subspace estimation in high dimensions

2012-11-02Unverified0· sign in to hype

Vincent Q. Vu, Jing Lei

Unverified — Be the first to reproduce this paper.

Reproduce

Abstract

We study sparse principal components analysis in high dimensions, where p (the number of variables) can be much larger than n (the number of observations), and analyze the problem of estimating the subspace spanned by the principal eigenvectors of the population covariance matrix. We introduce two complementary notions of _q subspace sparsity: row sparsity and column sparsity. We prove nonasymptotic lower and upper bounds on the minimax subspace estimation error for 0 q1. The bounds are optimal for row sparse subspaces and nearly optimal for column sparse subspaces, they apply to general classes of covariance matrices, and they show that _q constrained estimates can achieve optimal minimax rates without restrictive spiked covariance conditions. Interestingly, the form of the rates matches known results for sparse regression when the effective noise variance is defined appropriately. Our proof employs a novel variational theorem that may be useful in other regularized spectral estimation problems.

Tasks

Reproductions