Flexible sampling of discrete data correlations without the marginal distributions

06/12/2013
by   Alfredo Kalaitzis, et al.
0

Learning the joint dependence of discrete variables is a fundamental problem in machine learning, with many applications including prediction, clustering and dimensionality reduction. More recently, the framework of copula modeling has gained popularity due to its modular parametrization of joint distributions. Among other properties, copulas provide a recipe for combining flexible models for univariate marginal distributions with parametric families suitable for potentially high dimensional dependence structures. More radically, the extended rank likelihood approach of Hoff (2007) bypasses learning marginal models completely when such information is ancillary to the learning task at hand as in, e.g., standard dimensionality reduction problems or copula parameter estimation. The main idea is to represent data by their observable rank statistics, ignoring any other information from the marginals. Inference is typically done in a Bayesian framework with Gaussian copulas, and it is complicated by the fact this implies sampling within a space where the number of constraints increases quadratically with the number of data points. The result is slow mixing when using off-the-shelf Gibbs sampling. We present an efficient algorithm based on recent advances on constrained Hamiltonian Markov chain Monte Carlo that is simple to implement and does not require paying for a quadratic cost in sample size.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
12/19/2018

Pseudo-Marginal Hamiltonian Monte Carlo with Efficient Importance Sampling

The joint posterior of latent variables and parameters in Bayesian hiera...
research
11/11/2020

Solving high-dimensional parameter inference: marginal posterior densities Moment Networks

High-dimensional probability density estimation for inference suffers fr...
research
04/01/2018

Bayesian Mosaic: Parallelizable Composite Posterior

This paper proposes Bayesian mosaic, a parallelizable composite posterio...
research
06/06/2015

Fast Mixing for Discrete Point Processes

We investigate the systematic mechanism for designing fast mixing Markov...
research
08/26/2018

Bayesian inference for a single factor copula stochastic volatility model using Hamiltonian Monte Carlo

Single factor models are used in finance to model the joint behaviour of...
research
06/05/2019

A copula-based bivariate integer-valued autoregressive process with application

A bivariate integer-valued autoregressive process of order 1 (BINAR(1)) ...
research
04/08/2022

On Projectivity in Markov Logic Networks

Markov Logic Networks (MLNs) define a probability distribution on relati...

Please sign up or login with your details

Forgot password? Click here to reset