Higher-order interactions in statistical physics and machine learning: A non-parametric solution to the inverse problem

06/10/2020
by   Sjoerd Viktor Beentjes, et al.
0

We propose a model-independent definition of n-point interaction within a system of binary and categorical random variables from first principles, via the non-parametric framework of Targeted Learning, a subfield of mathematical statistics. This definition provides an interpretation for both magnitude and sign of 2-point, 3-point, and general n-point interactions. We show that the sign of an n-point interaction is interpretable relative to an (n-1)-point interaction obtained by fixing any one of the n variables. The non-parametric definition of interaction is fundamentally unbiased and reduces to familiar notions of interaction in parametric statistical physics models. Moreover, by taking into account information on conditional independence and without any further assumptions, the accuracy of interactions estimated directly from data is substantially increased whilst the number of samples required and the computational run time are both reduced. We illustrate these concepts both analytically and numerically on (i) the 2-dimensional Ising model, (ii) an Ising-like model with non-zero 2-point, 3-point, and 4-point interactions, (iii) the Restricted Boltzmann Machine (RBM), and argue that the formulation applies to energy-based models more generally. The non-parametric formulation allows for the direct reconstruction of the Hamiltonian from the data it generated. Finally, we discuss novel applications of this work, namely estimating causal molecular interactions leading to physiological outcomes, in population biomedicine.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
07/20/2018

Information Estimation Using Non-Parametric Copulas

Estimation of mutual information between random variables has become cru...
research
08/08/2017

Learning non-parametric Markov networks with mutual information

We propose a method for learning Markov network structures for continuou...
research
07/06/2020

Yield curve and macroeconomy interaction: evidence from the non-parametric functional lagged regression approach

Viewing a yield curve as a sparse collection of measurements on a latent...
research
06/01/2023

Interaction Measures, Partition Lattices and Kernel Tests for High-Order Interactions

Models that rely solely on pairwise relationships often fail to capture ...
research
12/11/2020

Online Coresets for Clustering with Bregman Divergences

We present algorithms that create coresets in an online setting for clus...
research
09/05/2023

Inferring effective couplings with Restricted Boltzmann Machines

Generative models offer a direct way to model complex data. Among them, ...
research
07/06/2019

Topological Information Data Analysis

This paper presents methods that quantify the structure of statistical i...

Please sign up or login with your details

Forgot password? Click here to reset