Learning Granularity-Aware Convolutional Neural Network for Fine-Grained Visual Classification

03/04/2021
by   Jianwei Song, et al.
0

Locating discriminative parts plays a key role in fine-grained visual classification due to the high similarities between different objects. Recent works based on convolutional neural networks utilize the feature maps taken from the last convolutional layer to mine discriminative regions. However, the last convolutional layer tends to focus on the whole object due to the large receptive field, which leads to a reduced ability to spot the differences. To address this issue, we propose a novel Granularity-Aware Convolutional Neural Network (GA-CNN) that progressively explores discriminative features. Specifically, GA-CNN utilizes the differences of the receptive fields at different layers to learn multi-granularity features, and it exploits larger granularity information based on the smaller granularity information found at the previous stages. To further boost the performance, we introduce an object-attentive module that can effectively localize the object given a raw image. GA-CNN does not need bounding boxes/part annotations and can be trained end-to-end. Extensive experimental results show that our approach achieves state-of-the-art performances on three benchmark datasets.

READ FULL TEXT

page 2

page 3

page 4

research
03/04/2021

Feature Boosting, Suppression, and Diversification for Fine-Grained Visual Classification

Learning feature representation from discriminative local regions plays ...
research
03/08/2020

Fine-Grained Visual Classification via Progressive Multi-Granularity Training of Jigsaw Patches

Fine-grained visual classification (FGVC) is much more challenging than ...
research
02/26/2019

Unsupervised Part Mining for Fine-grained Image Classification

Fine-grained image classification remains challenging due to the large i...
research
10/17/2022

Cross-layer Attention Network for Fine-grained Visual Categorization

Learning discriminative representations for subtle localized details pla...
research
09/02/2019

HiCoRe: Visual Hierarchical Context-Reasoning

Reasoning about images/objects and their hierarchical interactions is a ...
research
01/17/2021

Context-aware Attentional Pooling (CAP) for Fine-grained Visual Classification

Deep convolutional neural networks (CNNs) have shown a strong ability in...
research
05/15/2021

One for All: An End-to-End Compact Solution for Hand Gesture Recognition

The HGR is a quite challenging task as its performance is influenced by ...

Please sign up or login with your details

Forgot password? Click here to reset