Modelling the influence of data structure on learning in neural networks

09/25/2019
by   Sebastian Goldt, et al.
9

The lack of crisp mathematical models that capture the structure of real-world data sets is a major obstacle to the detailed theoretical understanding of deep neural networks. Here, we first demonstrate the effect of structured data sets by experimentally comparing the dynamics and the performance of two-layer networks trained on two different data sets: (i) an unstructured synthetic data set containing random i.i.d. inputs, and (ii) a simple canonical data set containing MNIST images. Our analysis reveals two phenomena related to the dynamics of the networks and their ability to generalise that only appear when training on structured data sets. Second, we introduce a generative model for data sets, where high-dimensional inputs lie on a lower-dimensional manifold and have labels that depend only on their position within this manifold. We call it the hidden manifold model and we experimentally demonstrate that training networks on data sets drawn from this model reproduces both the phenomena seen during training on MNIST.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/25/2020

The Gaussian equivalence of generative models for learning with two-layer neural networks

Understanding the impact of data structure on learning in neural network...
research
10/02/2020

Machine learning approach to force reconstruction in photoelastic materials

Photoelastic techniques have a long tradition in both qualitative and qu...
research
02/01/2023

Density peak clustering using tensor network

Tensor networks, which have been traditionally used to simulate many-bod...
research
09/27/2021

Abstraction, Reasoning and Deep Learning: A Study of the "Look and Say" Sequence

The ability to abstract, count, and use System 2 reasoning are well-know...
research
11/30/2020

Contagion Dynamics for Manifold Learning

Contagion maps exploit activation times in threshold contagions to assig...
research
08/29/2022

Dimension Independent Data Sets Approximation and Applications to Classification

We revisit the classical kernel method of approximation/interpolation th...
research
02/10/2017

Generative Mixture of Networks

A generative model based on training deep architectures is proposed. The...

Please sign up or login with your details

Forgot password? Click here to reset