Mixture Models, Robustness, and Sum of Squares Proofs

by   Samuel B. Hopkins, et al.

We use the Sum of Squares method to develop new efficient algorithms for learning well-separated mixtures of Gaussians and robust mean estimation, both in high dimensions, that substantially improve upon the statistical guarantees achieved by previous efficient algorithms. Firstly, we study mixtures of k distributions in d dimensions, where the means of every pair of distributions are separated by at least k^ε. In the special case of spherical Gaussian mixtures, we give a (dk)^O(1/ε^2)-time algorithm that learns the means assuming separation at least k^ε, for any ε > 0. This is the first algorithm to improve on greedy ("single-linkage") and spectral clustering, breaking a long-standing barrier for efficient algorithms at separation k^1/4. We also study robust estimation. When an unknown (1-ε)-fraction of X_1,...,X_n are chosen from a sub-Gaussian distribution with mean μ but the remaining points are chosen adversarially, we give an algorithm recovering μ to error ε^1-1/t in time d^O(t^2), so long as sub-Gaussian-ness up to O(t) moments can be certified by a Sum of Squares proof. This is the first polynomial-time algorithm with guarantees approaching the information-theoretic limit for non-Gaussian distributions. Previous algorithms could not achieve error better than ε^1/2. Both of these results are based on a unified technique. Inspired by recent algorithms of Diakonikolas et al. in robust statistics, we devise an SDP based on the Sum of Squares method for the following setting: given X_1,...,X_n ∈R^d for large d and n = poly(d) with the promise that a subset of X_1,...,X_n were sampled from a probability distribution with bounded moments, recover some information about that distribution.


page 1

page 2

page 3

page 4


Outlier-robust moment-estimation via sum-of-squares

We develop efficient algorithms for estimating low-degree moments of unk...

Beyond Parallel Pancakes: Quasi-Polynomial Time Guarantees for Non-Spherical Gaussian Mixtures

We consider mixtures of k≥ 2 Gaussian components with unknown means and ...

List-Decodable Robust Mean Estimation and Learning Mixtures of Spherical Gaussians

We study the problem of list-decodable Gaussian mean estimation and the ...

Outlier-Robust Clustering of Non-Spherical Mixtures

We give the first outlier-robust efficient algorithm for clustering a mi...

High-dimensional estimation via sum-of-squares proofs

Estimation is the computational task of recovering a hidden parameter x ...

Learning Mixtures of Gaussians Using the DDPM Objective

Recent works have shown that diffusion models can learn essentially any ...

On the robust learning mixtures of linear regressions

In this note, we consider the problem of robust learning mixtures of lin...

Please sign up or login with your details

Forgot password? Click here to reset