Skip to main navigation Skip to search Skip to main content

EXACT MINIMAX RISK for LINEAR LEAST SQUARES, and the LOWER TAIL of SAMPLE COVARIANCE MATRICES

  • ENSAE

Research output: Contribution to journalArticlepeer-review

29 Citations (Scopus)

Abstract

We consider random-design linear prediction and related questions on the lower tail of random matrices. It is known that, under boundedness constraints, the minimax risk is of order d/n in dimension d with n samples. Here, we study the minimax expected excess risk over the full linear class, depending on the distribution of covariates. First, the least squares estimator is exactly minimax optimal in the well-specified case, for every distribution of covariates. We express the minimax risk in terms of the distribution of statistical leverage scores of individual samples, and deduce a minimax lower bound of d/(n-d + 1) for any covariate distribution, nearly matching the risk for Gaussian design. We then obtain sharp nonasymptotic upper bounds for covariates that satisfy a "small ball"-Type regularity condition in both well-specified and misspecified cases. Our main technical contribution is the study of the lower tail of the smallest singular value of empirical covariance matrices at small values.We establish a lower bound on this lower tail, valid for any distribution in dimension d 2, together with a matching upper bound under a necessary regularity condition. Our proof relies on the PAC-Bayes technique for controlling empirical processes, and extends an analysis of Oliveira devoted to a different part of the lower tail.

Original languageEnglish
Pages (from-to)2157-2178
Number of pages22
JournalAnnals of Statistics
Volume50
Issue number4
DOIs
Publication statusPublished - 1 Aug 2022
Externally publishedYes

Keywords

  • Least squares
  • anticoncentration.
  • covariance matrices
  • decision theory
  • lower bounds
  • statistical learning theory

Fingerprint

Dive into the research topics of 'EXACT MINIMAX RISK for LINEAR LEAST SQUARES, and the LOWER TAIL of SAMPLE COVARIANCE MATRICES'. Together they form a unique fingerprint.

Cite this