SR-GNN: Spatial Relation-aware Graph Neural Network for Fine-Grained Image Categorization

09/05/2022
by   Asish Bera, et al.
28

Over the past few years, a significant progress has been made in deep convolutional neural networks (CNNs)-based image recognition. This is mainly due to the strong ability of such networks in mining discriminative object pose and parts information from texture and shape. This is often inappropriate for fine-grained visual classification (FGVC) since it exhibits high intra-class and low inter-class variances due to occlusions, deformation, illuminations, etc. Thus, an expressive feature representation describing global structural information is a key to characterize an object/ scene. To this end, we propose a method that effectively captures subtle changes by aggregating context-aware features from most relevant image-regions and their importance in discriminating fine-grained categories avoiding the bounding-box and/or distinguishable part annotations. Our approach is inspired by the recent advancement in self-attention and graph neural networks (GNNs) approaches to include a simple yet effective relation-aware feature transformation and its refinement using a context-aware attention mechanism to boost the discriminability of the transformed feature in an end-to-end learning process. Our model is evaluated on eight benchmark datasets consisting of fine-grained objects and human-object interactions. It outperforms the state-of-the-art approaches by a significant margin in recognition accuracy.

READ FULL TEXT

page 1

page 3

page 4

page 9

page 14

page 16

page 18

page 19

research
01/17/2021

Context-aware Attentional Pooling (CAP) for Fine-grained Visual Classification

Deep convolutional neural networks (CNNs) have shown a strong ability in...
research
10/23/2021

Attend and Guide (AG-Net): A Keypoints-driven Attention-based Deep Network for Image Recognition

This paper presents a novel keypoints-based attention mechanism for visu...
research
01/17/2021

Regional Attention Network (RAN) for Head Pose and Fine-grained Gesture Recognition

Affect is often expressed via non-verbal body language such as actions/g...
research
05/04/2022

Dual Cross-Attention Learning for Fine-Grained Visual Categorization and Object Re-Identification

Recently, self-attention mechanisms have shown impressive performance in...
research
03/23/2022

Spatial self-attention network with self-attention distillation for fine-grained image recognition

The underlining task for fine-grained image recognition captures both th...
research
10/20/2022

Towards Better Guided Attention and Human Knowledge Insertion in Deep Convolutional Neural Networks

Attention Branch Networks (ABNs) have been shown to simultaneously provi...
research
10/23/2021

An attention-driven hierarchical multi-scale representation for visual recognition

Convolutional Neural Networks (CNNs) have revolutionized the understandi...

Please sign up or login with your details

Forgot password? Click here to reset