Efficient Learning of Restricted Boltzmann Machines Using Covariance estimates

10/25/2018
by   Vidyadhar Upadhya, et al.
0

Learning of RBMs using standard algorithms such as CD(k) involves gradient descent on negative log-likelihood. One of the terms in the gradient, which is expectation of visible and hidden units is intractable and is obtained through an MCMC estimate. In this work we show that the Hessian of the log-likelihood can be written in terms of covariances of hidden and visible units and hence all elements of the Hessian can also be estimated using the same MCMC samples with minimal extra computational costs. Since inverting the Hessian may be computationally expensive, we propose an algorithm that uses inverse of the diagonal approximation of the Hessian. This essentially results in parameter-specific adaptive learning rates for the gradient descent process. We show that this algorithm improves the efficiency of learning RBMs compared to state-of-art methods. Specifically we show that using the inverse of diagonal approximation of Hessian in the stochastic DC (difference of convex functions) program approach results in very efficient learning of RBMs. We use different evaluation metrics to test the probability distribution learnt by the RBM along with the traditional criterion of average test and train log-likelihood.

READ FULL TEXT
research
02/11/2021

Learning Gaussian-Bernoulli RBMs using Difference of Convex Functions Optimization

The Gaussian-Bernoulli restricted Boltzmann machine (GB-RBM) is a useful...
research
09/21/2017

Learning RBM with a DC programming Approach

By exploiting the property that the RBM log-likelihood function is the d...
research
02/12/2015

Newton-based maximum likelihood estimation in nonlinear state space models

Maximum likelihood (ML) estimation using Newton's method in nonlinear st...
research
05/07/2020

Training and Classification using a Restricted Boltzmann Machine on the D-Wave 2000Q

Restricted Boltzmann Machine (RBM) is an energy based, undirected graphi...
research
02/23/2018

Accelerate iterated filtering

In simulation-based inferences for partially observed Markov process mod...
research
10/22/2020

Computationally and Statistically Efficient Truncated Regression

We provide a computationally and statistically efficient estimator for t...

Please sign up or login with your details

Forgot password? Click here to reset