Neural Optimal Control for Representation Learning

06/16/2020
by   Mathieu Chalvidal, et al.
4

The intriguing connections recently established between neural networks and dynamical systems have invited deep learning researchers to tap into the well-explored principles of differential calculus. Notably, the adjoint sensitivity method used in neural ordinary differential equations (Neural ODEs) has cast the training of neural networks as a control problem in which neural modules operate as continuous-time homeomorphic transformations of features. Typically, these methods optimize a single set of parameters governing the dynamical system for the whole data set, forcing the network to learn complex transformations that are functionally limited and computationally heavy. Instead, we propose learning a data-conditioned distribution of optimal controls over the network dynamics, emulating a form of input-dependent fast neural plasticity. We describe a general method for training such models as well as convergence proofs assuming mild hypotheses about the ODEs and show empirically that this method leads to simpler dynamics and reduces the computational cost of Neural ODEs. We evaluate this approach for unsupervised image representation learning; our new "functional" auto-encoding model with ODEs, AutoencODE, achieves state-of-the-art image reconstruction quality on CIFAR-10, and exhibits substantial improvements in unsupervised classification over existing auto-encoding models.

READ FULL TEXT

page 6

page 7

page 8

page 15

page 16

page 17

page 18

page 19

research
06/22/2022

Near-optimal control of dynamical systems with neural ordinary differential equations

Optimal control problems naturally arise in many scientific applications...
research
01/14/2022

Taylor-Lagrange Neural Ordinary Differential Equations: Toward Fast Training and Evaluation of Neural ODEs

Neural ordinary differential equations (NODEs) – parametrizations of dif...
research
10/20/2022

Neural ODEs as Feedback Policies for Nonlinear Optimal Control

Neural ordinary differential equations (Neural ODEs) model continuous ti...
research
12/16/2020

Physical deep learning based on optimal control of dynamical systems

A central topic in recent artificial intelligence technologies is deep l...
research
01/14/2019

AET vs. AED: Unsupervised Representation Learning by Auto-Encoding Transformations rather than Data

The success of deep neural networks often relies on a large amount of la...
research
04/11/2021

Weak Form Generalized Hamiltonian Learning

We present a method for learning generalized Hamiltonian decompositions ...
research
08/06/2020

Large-time asymptotics in deep learning

It is by now well-known that practical deep supervised learning may roug...

Please sign up or login with your details

Forgot password? Click here to reset