Prediction when fitting simple models to high-dimensional data 1

2Citations
Citations of this article
10Readers
Mendeley users who have this article in their library.

Abstract

We study linear subset regression in the context of a high-dimensional linear model. Consider y = ϑ + θz + with univariate response y and a dvector of random regressors z, and a submodel where y is regressed on a set of p explanatory variables that are given by x = Mz, for some d × p matrix M. Here, “high-dimensional” means that the number d of available explanatory variables in the overall model is much larger than the number p of variables in the submodel. In this paper, we present Pinsker-type results for prediction of y given x. In particular, we show that the mean squared prediction error of the best linear predictor of y given x is close to the mean squared prediction error of the corresponding Bayes predictor E[yx], provided only that p/log d is small. We also show that the mean squared prediction error of the (feasible) least-squares predictor computed from n independent observations of (y, x) is close to that of the Bayes predictor, provided only that both p/log d and p/n are small. Our results hold uniformly in the regression parameters and over large collections of distributions for the design variables z.

Cite

CITATION STYLE

APA

Steinberger, L., & Leeb, H. (2019). Prediction when fitting simple models to high-dimensional data 1. Annals of Statistics, 47(3), 1408–1442. https://doi.org/10.1214/18-AOS1719

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free