Towards Real-Time Action Recognition on Mobile Devices Using Deep Models

06/17/2019
by   Chen-Lin Zhang, et al.
0

Action recognition is a vital task in computer vision, and many methods are developed to push it to the limit. However, current action recognition models have huge computational costs, which cannot be deployed to real-world tasks on mobile devices. In this paper, we first illustrate the setting of real-time action recognition, which is different from current action recognition inference settings. Under the new inference setting, we investigate state-of-the-art action recognition models on the Kinetics dataset empirically. Our results show that designing efficient real-time action recognition models is different from designing efficient ImageNet models, especially in weight initialization. We show that pre-trained weights on ImageNet improve the accuracy under the real-time action recognition setting. Finally, we use the hand gesture recognition task as a case study to evaluate our compact real-time action recognition models in real-world applications on mobile phones. Results show that our action recognition models, being 6x faster and with similar accuracy as state-of-the-art, can roughly meet the real-time requirements on mobile devices. To our best knowledge, this is the first paper that deploys current deep learning action recognition models on mobile devices.

READ FULL TEXT
research
08/27/2019

Mobile Video Action Recognition

Video action recognition, which is topical in computer vision and video ...
research
06/06/2018

Action4D: Real-time Action Recognition in the Crowd and Clutter

Recognizing every person's action in a crowded and cluttered environment...
research
10/27/2018

Real-time Action Recognition with Dissimilarity-based Training of Specialized Module Networks

This paper addresses the problem of real-time action recognition in trim...
research
05/21/2019

Lightweight Network Architecture for Real-Time Action Recognition

In this work we present a new efficient approach to Human Action Recogni...
research
10/15/2019

Real-time monitoring of driver drowsiness on mobile platforms using 3D neural networks

Driver drowsiness increases crash risk, leading to substantial road trau...
research
02/22/2022

ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing Users

Recent advances have enabled automatic sound recognition systems for dea...
research
12/29/2017

Learning Deep and Compact Models for Gesture Recognition

We look at the problem of developing a compact and accurate model for ge...

Please sign up or login with your details

Forgot password? Click here to reset