Multi-Fiber Networks for Video Recognition

07/30/2018
by   Yunpeng Chen, et al.
0

In this paper, we aim to reduce the computational cost of spatio-temporal deep neural networks, making them run as fast as their 2D counterparts while preserving state-of-the-art accuracy on video recognition benchmarks. To this end, we present the novel Multi-Fiber architecture that slices a complex neural network into an ensemble of lightweight networks or fibers that run through the network. To facilitate information flow between fibers we further incorporate multiplexer modules and end up with an architecture that reduces the computational cost of 3D networks by an order of magnitude, while increasing recognition performance at the same time. Extensive experimental results show that our multi-fiber architecture significantly boosts the efficiency of existing convolution networks for both image and video recognition tasks, achieving state-of-the-art performance on UCF-101, HMDB-51 and Kinetics datasets. Our proposed model requires over 9x and 13x less computations than the I3D and R(2+1)D models, respectively, yet providing higher accuracy.

READ FULL TEXT
research
12/05/2021

STSM: Spatio-Temporal Shift Module for Efficient Action Recognition

The modeling, computational cost, and accuracy of traditional Spatio-tem...
research
11/07/2016

Sigma Delta Quantized Networks

Deep neural networks can be obscenely wasteful. When processing video, a...
research
05/23/2022

Online Hybrid Lightweight Representations Learning: Its Application to Visual Tracking

This paper presents a novel hybrid representation learning framework for...
research
07/10/2019

Video Action Recognition Via Neural Architecture Searching

Deep neural networks have achieved great success for video analysis and ...
research
05/31/2019

Design Light-weight 3D Convolutional Networks for Video Recognition Temporal Residual, Fully Separable Block, and Fast Algorithm

Deep 3-dimensional (3D) Convolutional Network (ConvNet) has shown promis...
research
05/24/2022

Semi-Parametric Deep Neural Networks in Linear Time and Memory

Recent advances in deep learning have been driven by large-scale paramet...
research
10/27/2018

A^2-Nets: Double Attention Networks

Learning to capture long-range relations is fundamental to image/video r...

Please sign up or login with your details

Forgot password? Click here to reset