Generalization Through the Lens of Learning Dynamics

12/11/2022
by   Clare Lyle, et al.
0

A machine learning (ML) system must learn not only to match the output of a target function on a training set, but also to generalize to novel situations in order to yield accurate predictions at deployment. In most practical applications, the user cannot exhaustively enumerate every possible input to the model; strong generalization performance is therefore crucial to the development of ML systems which are performant and reliable enough to be deployed in the real world. While generalization is well-understood theoretically in a number of hypothesis classes, the impressive generalization performance of deep neural networks has stymied theoreticians. In deep reinforcement learning (RL), our understanding of generalization is further complicated by the conflict between generalization and stability in widely-used RL algorithms. This thesis will provide insight into generalization by studying the learning dynamics of deep neural networks in both supervised and reinforcement learning tasks.

READ FULL TEXT
research
11/30/2018

An Introduction to Deep Reinforcement Learning

Deep reinforcement learning is the combination of reinforcement learning...
research
08/03/2020

Dynamics Generalization via Information Bottleneck in Deep Reinforcement Learning

Despite the significant progress of deep reinforcement learning (RL) in ...
research
03/15/2022

Towards understanding deep learning with the natural clustering prior

The prior knowledge (a.k.a. priors) integrated into the design of a mach...
research
09/22/2020

The Relativity of Induction

Lately there has been a lot of discussion about why deep learning algori...
research
09/29/2018

Generalization and Regularization in DQN

Deep reinforcement learning (RL) algorithms have shown an impressive abi...
research
03/02/2023

Understanding plasticity in neural networks

Plasticity, the ability of a neural network to quickly change its predic...
research
04/27/2023

Learning to Extrapolate: A Transductive Approach

Machine learning systems, especially with overparameterized deep neural ...

Please sign up or login with your details

Forgot password? Click here to reset