On the Symmetries of Deep Learning Models and their Internal Representations

05/27/2022
by   Charles Godfrey, et al.
0

Symmetry has been a fundamental tool in the exploration of a broad range of complex systems. In machine learning, symmetry has been explored in both models and data. In this paper we seek to connect the symmetries arising from the architecture of a family of models with the symmetries of that family's internal representation of data. We do this by calculating a set of fundamental symmetry groups, which we call the intertwiner groups of the model. Each of these arises from a particular nonlinear layer of the model and different nonlinearities result in different symmetry groups. These groups change the weights of a model in such a way that the underlying function that the model represents remains constant but the internal representations of data inside the model may change. We connect intertwiner groups to a model's internal representations of data through a range of experiments that probe similarities between hidden states across models with the same architecture. Our work suggests that the symmetries of a network are propagated into the symmetries in that network's representation of data, providing us with a better understanding of how architecture affects the learning and prediction process. Finally, we speculate that for ReLU networks, the intertwiner groups may provide a justification for the common practice of concentrating model interpretability exploration on the activation basis in hidden layers rather than arbitrary linear combinations thereof.

READ FULL TEXT

page 3

page 8

page 18

page 19

page 30

page 31

page 32

page 33

research
04/15/2010

Symmetry within Solutions

We define the concept of an internal symmetry. This is a symmety within ...
research
12/20/2014

Classifier with Hierarchical Topographical Maps as Internal Representation

In this study we want to connect our previously proposed context-relevan...
research
02/25/2023

Two-Disk Compound Symmetry Groups

Symmetry is at the heart of much of mathematics, physics, and art. Tradi...
research
02/22/2021

Wallpaper group kirigami

Kirigami, the art of paper cutting, has become a paradigm for mechanical...
research
08/03/2023

A Novel Convolutional Neural Network Architecture with a Continuous Symmetry

This paper introduces a new Convolutional Neural Network (ConvNet) archi...
research
10/12/2017

Analysis of planar ornament patterns via motif asymmetry assumption and local connections

Planar ornaments, a.k.a. wallpapers, are regular repetitive patterns whi...
research
02/22/2023

Stress and Adaptation: Applying Anna Karenina Principle in Deep Learning for Image Classification

Image classification with deep neural networks has reached state-of-art ...

Please sign up or login with your details

Forgot password? Click here to reset