List-Decodable Covariance Estimation

06/22/2022
by   Misha Ivkov, et al.
0

We give the first polynomial time algorithm for list-decodable covariance estimation. For any α > 0, our algorithm takes input a sample Y ⊆ℝ^d of size n≥ d^𝗉𝗈𝗅𝗒(1/α) obtained by adversarially corrupting an (1-α)n points in an i.i.d. sample X of size n from the Gaussian distribution with unknown mean μ_* and covariance Σ_*. In n^𝗉𝗈𝗅𝗒(1/α) time, it outputs a constant-size list of k = k(α)= (1/α)^𝗉𝗈𝗅𝗒(1/α) candidate parameters that, with high probability, contains a (μ̂,Σ̂) such that the total variation distance TV(𝒩(μ_*,Σ_*),𝒩(μ̂,Σ̂))<1-O_α(1). This is the statistically strongest notion of distance and implies multiplicative spectral and relative Frobenius distance approximation for parameters with dimension independent error. Our algorithm works more generally for (1-α)-corruptions of any distribution D that possesses low-degree sum-of-squares certificates of two natural analytic properties: 1) anti-concentration of one-dimensional marginals and 2) hypercontractivity of degree 2 polynomials. Prior to our work, the only known results for estimating covariance in the list-decodable setting were for the special cases of list-decodable linear regression and subspace recovery due to Karmarkar, Klivans, and Kothari (2019), Raghavendra and Yau (2019 and 2020) and Bakshi and Kothari (2020). These results need superpolynomial time for obtaining any subconstant error in the underlying dimension. Our result implies the first polynomial-time exact algorithm for list-decodable linear regression and subspace recovery that allows, in particular, to obtain 2^-𝗉𝗈𝗅𝗒(d) error in polynomial-time. Our result also implies an improved algorithm for clustering non-spherical mixtures.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/14/2019

List-Decodable Linear Regression

We give the first polynomial-time algorithm for robust regression in the...
research
11/20/2017

List-Decodable Robust Mean Estimation and Learning Mixtures of Spherical Gaussians

We study the problem of list-decodable Gaussian mean estimation and the ...
research
02/12/2020

List-Decodable Subspace Recovery via Sum-of-Squares

We give the first efficient algorithm for the problem of list-decodable ...
research
05/26/2022

On Learning Mixture of Linear Regressions in the Non-Realizable Setting

While mixture of linear regressions (MLR) is a well-studied topic, prior...
research
05/06/2020

Outlier-Robust Clustering of Non-Spherical Mixtures

We give the first outlier-robust efficient algorithm for clustering a mi...
research
06/16/2021

Breaking The Dimension Dependence in Sparse Distribution Estimation under Communication Constraints

We consider the problem of estimating a d-dimensional s-sparse discrete ...
research
11/19/2020

List-Decodable Mean Estimation in Nearly-PCA Time

Traditionally, robust statistics has focused on designing estimators tol...

Please sign up or login with your details

Forgot password? Click here to reset