A fast algorithm for computing distance correlation

10/26/2018
by   Arin Chaudhuri, et al.
0

Classical dependence measures such as Pearson correlation, Spearman's ρ, and Kendall's τ can detect only monotonic or linear dependence. To overcome these limitations, szekely2007measuring proposed distance covariance as a weighted L_2 distance between the joint characteristic function and the product of marginal distributions. The distance covariance is 0 if and only if two random vectors X and Y are independent. This measure has the power to detect the presence of a dependence structure when the sample size is large enough. They further showed that the sample distance covariance can be calculated simply from modified Euclidean distances, which typically requires O(n^2) cost. The quadratic computing time greatly limits the application of distance covariance to large data. In this paper, we present a simple exact O(n(n)) algorithm to calculate the sample distance covariance between two univariate random variables. The proposed method essentially consists of two sorting steps, so it is easy to implement. Empirical results show that the proposed algorithm is significantly faster than state-of-the-art methods. The algorithm's speed will enable researchers to explore complicated dependence structures in large datasets.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
11/21/2017

Detecting independence of random vectors II. Distance multivariance and Gaussian multivariance

We introduce two new measures for the dependence of n > 2 random variabl...
research
11/25/2017

Distance Metrics for Measuring Joint Dependence with Application to Causal Inference

Many statistical applications require the quantification of joint depend...
research
06/05/2019

Estimating Feature-Label Dependence Using Gini Distance Statistics

Identifying statistical dependence between the features and the label is...
research
06/01/2015

Mutual Dependence: A Novel Method for Computing Dependencies Between Random Variables

In data science, it is often required to estimate dependencies between d...
research
06/27/2012

Learning Markov Network Structure using Brownian Distance Covariance

In this paper, we present a simple non-parametric method for learning th...
research
02/17/2021

Estimating The Proportion of Signal Variables Under Arbitrary Covariance Dependence

Estimating the proportion of signals hidden in a large amount of noise v...
research
06/21/2022

A Basic Treatment of the Distance Covariance

The distance covariance of Székely, et al. [23] and Székely and Rizzo [2...

Please sign up or login with your details

Forgot password? Click here to reset