Robust score matching for compositional data

05/12/2023
by   Janice L. Scealy, et al.
0

The restricted polynomially-tilted pairwise interaction (RPPI) distribution gives a flexible model for compositional data. It is particularly well-suited to situations where some of the marginal distributions of the components of a composition are concentrated near zero, possibly with right skewness. This article develops a method of tractable robust estimation for the model by combining two ideas. The first idea is to use score matching estimation after an additive log-ratio transformation. The resulting estimator is automatically insensitive to zeros in the data compositions. The second idea is to incorporate suitable weights in the estimating equations. The resulting estimator is additionally resistant to outliers. These properties are confirmed in simulation studies where we further also demonstrate that our new outlier-robust estimator is efficient in high concentration settings, even in the case when there is no model contamination. An example is given using microbiome data. A user-friendly R package accompanies the article.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
12/23/2020

Score matching for compositional distributions

Compositional data and multivariate count data with known totals are cha...
research
06/08/2022

Robust self-tuning semiparametric PCA for contaminated elliptical distribution

Principal component analysis (PCA) is one of the most popular dimension ...
research
02/20/2018

A folded model for compositional data analysis

A folded type model is developed for analyzing compositional data. The p...
research
06/27/2023

Triply robust estimation under missing at random

Missing data is frequently encountered in many areas of statistics. Impu...
research
09/10/2021

Interaction Models and Generalized Score Matching for Compositional Data

Applications such as the analysis of microbiome data have led to renewed...
research
06/08/2018

Estimation of marginal model with subgroup auxiliary information

Marginal model is a popular instrument for studying longitudinal data an...
research
09/13/2023

CARE: Large Precision Matrix Estimation for Compositional Data

High-dimensional compositional data are prevalent in many applications. ...

Please sign up or login with your details

Forgot password? Click here to reset