Searching a High-Performance Feature Extractor for Text Recognition Network

09/27/2022
by   Hui Zhang, et al.
0

Feature extractor plays a critical role in text recognition (TR), but customizing its architecture is relatively less explored due to expensive manual tweaking. In this work, inspired by the success of neural architecture search (NAS), we propose to search for suitable feature extractors. We design a domain-specific search space by exploring principles for having good feature extractors. The space includes a 3D-structured space for the spatial model and a transformed-based space for the sequential model. As the space is huge and complexly structured, no existing NAS algorithms can be applied. We propose a two-stage algorithm to effectively search in the space. In the first stage, we cut the space into several blocks and progressively train each block with the help of an auxiliary head. We introduce the latency constraint into the second stage and search sub-network from the trained supernet via natural gradient descent. In experiments, a series of ablation studies are performed to better understand the designed space, search algorithm, and searched architectures. We also compare the proposed method with various state-of-the-art ones on both hand-written and scene TR tasks. Extensive results show that our approach can achieve better recognition performance with less latency.

READ FULL TEXT
research
03/14/2020

Efficient Backbone Search for Scene Text Recognition

Scene text recognition (STR) is very challenging due to the diversity of...
research
03/13/2022

Training Protocol Matters: Towards Accurate Scene Text Recognition via Training Protocol Searching

The development of scene text recognition (STR) in the era of deep learn...
research
09/08/2021

RepNAS: Searching for Efficient Re-parameterizing Blocks

In the past years, significant improvements in the field of neural archi...
research
06/10/2020

AMER: Automatic Behavior Modeling and Interaction Exploration in Recommender System

User behavior and feature interactions are crucial in deep learning-base...
research
01/24/2022

Neural Architecture Searching for Facial Attributes-based Depression Recognition

Recent studies show that depression can be partially reflected from huma...
research
03/26/2022

AutoTS: Automatic Time Series Forecasting Model Design Based on Two-Stage Pruning

Automatic Time Series Forecasting (TSF) model design which aims to help ...
research
07/23/2020

Representation Sharing for Fast Object Detector Search and Beyond

Region Proposal Network (RPN) provides strong support for handling the s...

Please sign up or login with your details

Forgot password? Click here to reset