Thermodynamics of Restricted Boltzmann Machines and related learning dynamics

03/05/2018
by   Aurelien Decelle, et al.
0

We analyze the learning process of the restricted Boltzmann machine (RBM), a certain type of generative models used in the context of unsupervised learning. In a first step, we investigate the thermodynamics properties by considering a realistic statistical ensemble of RBM, assuming the information content of the RBM to be mainly reflected by spectral properties of its weight matrix W. A phase diagram is obtained which seems at first sight similar to the one of the Sherrington-Kirkpatrick (SK) model with ferromagnetic couplings. The main difference resides in the structure of the ferromagnetic phase which may or may not be of compositional type, depending mainly on the distribution's kurtosis of the singular vectors components of W. In a second step the learning dynamics of an RBM from arbitrary data is studied in thermodynamic limit. A "typical" learning trajectory is shown to solve an effective dynamical equation, based on the aforementioned ensemble average and involving explicitly order parameters obtained from the thermodynamic analysis. This accounts in particular for the dominant singular values evolution and how this is driven by the input data: in the linear regime at the beginning of the learning, they correspond to unstable deformation modes of W reflecting dominant covariance modes of the data. In the non-linear regime it is seen how the selected modes interact in later stages of the learning procedure, by eventually imposing a matching between order parameters with their empirical counterparts estimated from the data. Experiments on both artificial and real data illustrate these considerations, showing in particular how the RBM operates in the ferromagnetic compositional phase.

READ FULL TEXT
research
11/23/2020

Restricted Boltzmann Machine, recent advances and mean-field theory

This review deals with Restricted Boltzmann Machine (RBM) under the ligh...
research
10/31/2019

Gaussian-Spherical Restricted Boltzmann Machines

We consider a special type of Restricted Boltzmann machine (RBM), namely...
research
11/21/2016

Emergence of Compositional Representations in Restricted Boltzmann Machines

Extracting automatically the complex set of features composing real high...
research
04/25/2018

Improved Classification Based on Deep Belief Networks

For better classification generative models are used to initialize the m...
research
05/04/2020

A Dynamical Mean-Field Theory for Learning in Restricted Boltzmann Machines

We define a message-passing algorithm for computing magnetizations in Re...
research
06/27/2012

A Generative Process for Sampling Contractive Auto-Encoders

The contractive auto-encoder learns a representation of the input data t...
research
05/11/2023

Investigating the generative dynamics of energy-based neural networks

Generative neural networks can produce data samples according to the sta...

Please sign up or login with your details

Forgot password? Click here to reset