Data-driven regularization of Wasserstein barycenters with an application to multivariate density registration

04/24/2018
by   Jérémie Bigot, et al.
0

We present a framework to simultaneously align and smooth data in the form of multiple point clouds sampled from unknown densities with support in a d-dimensional Euclidean space. This work is motivated by applications in bio-informatics where researchers aim to automatically normalize large datasets to compare and analyze characteristics within a same cell population. Inconveniently, the information acquired is noisy due to mis-alignment caused by technical variations of the environment. To overcome this problem, we propose to register multiple point clouds by using the notion of regularized barycenter (or Fréchet mean) of a set of probability measures with respect to the Wasserstein metric which allows to smooth such data and to remove mis-alignment effect in the sample acquisition process. A first approach consists in penalizing a Wasserstein barycenter with a convex functional as recently proposed in Bigot and al. (2018). A second strategy is to modify the Wasserstein metric itself by using an entropically regularized transportation cost between probability measures as introduced in Cuturi (2013). The main contribution of this work is to propound data-driven choices for the regularization parameters involved in each approach using the Goldenshluger-Lepski's principle. Simulated data sampled from Gaussian mixtures are used to illustrate each method, and an application to the analysis of flow cytometry data is finally proposed.

READ FULL TEXT

page 18

page 19

page 20

research
03/09/2015

A Smoothed Dual Approach for Variational Wasserstein Problems

Variational problems that involve Wasserstein distances have been recent...
research
07/25/2023

Computational Guarantees for Doubly Entropic Wasserstein Barycenters via Damped Sinkhorn Iterations

We study the computation of doubly regularized Wasserstein barycenters, ...
research
03/21/2023

Doubly Regularized Entropic Wasserstein Barycenters

We study a general formulation of regularized Wasserstein barycenters th...
research
06/16/2020

CytOpT: Optimal Transport with Domain Adaptation for Interpreting Flow Cytometry data

The automated analysis of flow cytometry measurements is an active resea...
research
06/17/2015

Learning with a Wasserstein Loss

Learning to predict multi-label outputs is challenging, but in many prob...
research
10/25/2022

Wasserstein Archetypal Analysis

Archetypal analysis is an unsupervised machine learning method that summ...
research
08/25/2022

A deep learning framework for geodesics under spherical Wasserstein-Fisher-Rao metric and its application for weighted sample generation

Wasserstein-Fisher-Rao (WFR) distance is a family of metrics to gauge th...

Please sign up or login with your details

Forgot password? Click here to reset