Enabling Deep Learning on Edge Devices through Filter Pruning and Knowledge Transfer

01/22/2022
by   Kaiqi Zhao, et al.
0

Deep learning models have introduced various intelligent applications to edge devices, such as image classification, speech recognition, and augmented reality. There is an increasing need of training such models on the devices in order to deliver personalized, responsive, and private learning. To address this need, this paper presents a new solution for deploying and training state-of-the-art models on the resource-constrained devices. First, the paper proposes a novel filter-pruning-based model compression method to create lightweight trainable models from large models trained in the cloud, without much loss of accuracy. Second, it proposes a novel knowledge transfer method to enable the on-device model to update incrementally in real time or near real time using incremental learning on new data and enable the on-device model to learn the unseen categories with the help of the in-cloud model in an unsupervised fashion. The results show that 1) our model compression method can remove up to 99.36 of over 90 compressed models to achieve more than 90 accuracy on old categories; 3) it allows the compressed models to converge within real time (three to six minutes) on the edge for incremental learning tasks; 4) it enables the model to classify unseen categories of data (78.92 Top-1 accuracy) that it is never trained with.

READ FULL TEXT
research
03/25/2021

Real-time low-resource phoneme recognition on edge devices

While speech recognition has seen a surge in interest and research over ...
research
08/03/2019

Real-time Deep Learning at the Edge for Scalable Reliability Modeling of Si-MOSFET Power Electronics Converters

With the significant growth of advanced high-frequency power converters,...
research
08/09/2023

FPGA Resource-aware Structured Pruning for Real-Time Neural Networks

Neural networks achieve state-of-the-art performance in image classifica...
research
02/09/2021

Sparsification via Compressed Sensing for Automatic Speech Recognition

In order to achieve high accuracy for machine learning (ML) applications...
research
03/25/2021

Prototype-based Personalized Pruning

Nowadays, as edge devices such as smartphones become prevalent, there ar...
research
11/20/2021

Real-time Human Detection Model for Edge Devices

Building a small-sized fast surveillance system model to fit on limited ...
research
03/05/2021

Environmental Sound Classification on the Edge: A Pipeline for Deep Acoustic Networks on Extremely Resource-Constrained Devices

Significant efforts are being invested to bring state-of-the-art classif...

Please sign up or login with your details

Forgot password? Click here to reset