Outlier-Robust Optimal Transport: Duality, Structure, and Statistical Analysis

11/02/2021
by   Sloan Nietert, et al.
0

The Wasserstein distance, rooted in optimal transport (OT) theory, is a popular discrepancy measure between probability distributions with various applications to statistics and machine learning. Despite their rich structure and demonstrated utility, Wasserstein distances are sensitive to outliers in the considered distributions, which hinders applicability in practice. Inspired by the Huber contamination model, we propose a new outlier-robust Wasserstein distance 𝖶_p^ε which allows for ε outlier mass to be removed from each contaminated distribution. Our formulation amounts to a highly regular optimization problem that lends itself better for analysis compared to previously considered frameworks. Leveraging this, we conduct a thorough theoretical study of 𝖶_p^ε, encompassing characterization of optimal perturbations, regularity, duality, and statistical estimation and robustness results. In particular, by decoupling the optimization variables, we arrive at a simple dual form for 𝖶_p^ε that can be implemented via an elementary modification to standard, duality-based OT solvers. We illustrate the benefits of our framework via applications to generative modeling with contaminated datasets.

READ FULL TEXT

page 9

page 13

research
02/02/2023

Robust Estimation under the Wasserstein Distance

We study the problem of robust distribution estimation under the Wassers...
research
12/14/2020

Outlier-Robust Optimal Transport

Optimal transport (OT) provides a way of measuring distances between dis...
research
04/30/2022

A Simple Duality Proof for Wasserstein Distributionally Robust Optimization

We present a short and elementary proof of the duality for Wasserstein d...
research
09/03/2022

Entropy-regularized Wasserstein distributionally robust shape and topology optimization

This brief note aims to introduce the recent paradigm of distributional ...
research
09/08/2019

Distributionally Robust Optimization with Correlated Data from Vector Autoregressive Processes

We present a distributionally robust formulation of a stochastic optimiz...
research
01/16/2023

Theoretical and computational aspects of robust optimal transportation, with applications to statistics and machine learning

Optimal transport (OT) theory and the related p-Wasserstein distance (W_...
research
12/25/2022

Gromov-Wasserstein Distances: Entropic Regularization, Duality, and Sample Complexity

The Gromov-Wasserstein (GW) distance quantifies dissimilarity between me...

Please sign up or login with your details

Forgot password? Click here to reset