End-To-End Data-Dependent Routing in Multi-Path Neural Networks

07/06/2021
by   Dumindu Tissera, et al.
0

Neural networks are known to give better performance with increased depth due to their ability to learn more abstract features. Although the deepening of networks has been well established, there is still room for efficient feature extraction within a layer which would reduce the need for mere parameter increment. The conventional widening of networks by having more filters in each layer introduces a quadratic increment of parameters. Having multiple parallel convolutional/dense operations in each layer solves this problem, but without any context-dependent allocation of resources among these operations: the parallel computations tend to learn similar features making the widening process less effective. Therefore, we propose the use of multi-path neural networks with data-dependent resource allocation among parallel computations within layers, which also lets an input to be routed end-to-end through these parallel paths. To do this, we first introduce a cross-prediction based algorithm between parallel tensors of subsequent layers. Second, we further reduce the routing overhead by introducing feature-dependent cross-connections between parallel tensors of successive layers. Our multi-path networks show superior performance to existing widening and adaptive feature extraction, and even ensembles, and deeper networks at similar complexity in the image recognition task.

READ FULL TEXT

page 2

page 13

research
06/24/2020

Feature-dependent Cross-Connections in Multi-Path Neural Networks

Learning a particular task from a dataset, samples in which originate fr...
research
07/26/2019

Context-Aware Multipath Networks

Making a single network effectively address diverse contexts---learning ...
research
11/17/2016

DelugeNets: Deep Networks with Efficient and Flexible Cross-layer Information Inflows

Deluge Networks (DelugeNets) are deep neural networks which efficiently ...
research
08/01/2017

CREST: Convolutional Residual Learning for Visual Tracking

Discriminative correlation filters (DCFs) have been shown to perform sup...
research
07/09/2021

Multi-path Convolutional Neural Networks Efficiently Improve Feature Extraction in Continuous Adventitious Lung Sound Detection

We previously established a large lung sound database, HF_Lung_V2 (Lung_...
research
01/19/2019

Towards Universal End-to-End Affect Recognition from Multilingual Speech by ConvNets

We propose an end-to-end affect recognition approach using a Convolution...
research
03/01/2023

Feature Extraction Matters More: Universal Deepfake Disruption through Attacking Ensemble Feature Extractors

Adversarial example is a rising way of protecting facial privacy securit...

Please sign up or login with your details

Forgot password? Click here to reset