A survey on modern trainable activation functions

05/02/2020
by   Andrea Apicella, et al.
0

In the literature, there is a strong interest to identify and define activation functions which can improve neural network performance. In recent years there is a renovated interest of the scientific community in investigating activation functions which can be trained during the learning process, usually referred as trainable, learnable or adaptable activation functions. They appear to lead to better network performance. Diverse and heterogeneous models of trainable activation function have been proposed in the literature. In this paper, we present a survey of these models. Starting from a discussion on the use of the term "activation function" in literature, we propose a taxonomy of trainable activation functions, highlight common and distinctive proprieties of recent and past models, and discuss on main advantages and limitations of this type of approach. We show that many of the proposed approaches are equivalent to add neuron layers which use fixed activation functions (nontrainable activation functions) and some simple local rule constrains the corresponding weight layers.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/02/2023

ErfReLU: Adaptive Activation Function for Deep Neural Network

Recent research has found that the activation function (AF) selected for...
research
09/06/2022

How important are activation functions in regression and classification? A survey, performance comparison, and future directions

Inspired by biological neurons, the activation functions play an essenti...
research
08/30/2022

Transformers with Learnable Activation Functions

Activation functions can have a significant impact on reducing the topol...
research
08/07/2022

Transmission Neural Networks: From Virus Spread Models to Neural Networks

This work connects models for virus spread on networks with their equiva...
research
11/29/2017

Gaussian Process Neurons Learn Stochastic Activation Functions

We propose stochastic, non-parametric activation functions that are full...
research
05/18/2016

Learning activation functions from data using cubic spline interpolation

Neural networks require a careful design in order to perform properly on...

Please sign up or login with your details

Forgot password? Click here to reset