A zero-estimator approach for estimating the signal level in a   high-dimensional regression setting

Ilan Livne

arXiv:2307.06739·math.ST·July 26, 2023·1 cites

A zero-estimator approach for estimating the signal level in a high-dimensional regression setting

Ilan Livne

PDF

Open Access

TL;DR

This paper introduces a zero-estimator approach to improve the estimation of signal and noise levels in high-dimensional regression, especially leveraging unlabeled data to enhance accuracy.

Contribution

It proposes a novel zero-estimator method that improves naive estimators for signal and noise levels in high-dimensional semi-supervised regression models.

Findings

01

Zero-estimator improves variance reduction in estimators.

02

Method performs well on four real datasets.

03

Approach enhances estimation accuracy using unlabeled data.

Abstract

Analysis of high-dimensional data, where the number of covariates is larger than the sample size, is a topic of current interest. In such settings, an important goal is to estimate the signal level $τ^{2}$ and noise level $σ^{2}$ , i.e., to quantify how much variation in the response variable can be explained by the covariates, versus how much of the variation is left unexplained. This thesis considers the estimation of these quantities in a semi-supervised setting, where for many observations only the vector of covariates $X$ is given with no responses $Y$ . Our main research question is: how can one use the unlabeled data to better estimate $τ^{2}$ and $σ^{2}$ ? We consider two frameworks: a linear regression model and a linear projection model in which linearity is not assumed. In the first framework, while linear regression is used, no sparsity assumptions on the coefficients…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsStatistical Methods and Inference · Advanced Statistical Methods and Models · Advanced Statistical Process Monitoring