A Novel Plug-in Module for Fine-Grained Visual Classification

02/08/2022
by   Po-Yung Chou, et al.
0

Visual classification can be divided into coarse-grained and fine-grained classification. Coarse-grained classification represents categories with a large degree of dissimilarity, such as the classification of cats and dogs, while fine-grained classification represents classifications with a large degree of similarity, such as cat species, bird species, and the makes or models of vehicles. Unlike coarse-grained visual classification, fine-grained visual classification often requires professional experts to label data, which makes data more expensive. To meet this challenge, many approaches propose to automatically find the most discriminative regions and use local features to provide more precise features. These approaches only require image-level annotations, thereby reducing the cost of annotation. However, most of these methods require two- or multi-stage architectures and cannot be trained end-to-end. Therefore, we propose a novel plug-in module that can be integrated to many common backbones, including CNN-based or Transformer-based networks to provide strongly discriminative regions. The plugin module can output pixel-level feature maps and fuse filtered features to enhance fine-grained visual classification. Experimental results show that the proposed plugin module outperforms state-of-the-art approaches and significantly improves the accuracy to 92.77% and 92.83% on CUB200-2011 and NABirds, respectively. We have released our source code in Github https://github.com/chou141253/FGVC-PIM.git.

READ FULL TEXT

page 1

page 8

research
04/18/2020

Feathers dataset for Fine-Grained Visual Categorization

This paper introduces a novel dataset FeatherV1, containing 28,272 image...
research
08/31/2023

Coarse-to-Fine Amodal Segmentation with Shape Prior

Amodal object segmentation is a challenging task that involves segmentin...
research
07/21/2021

Automated Refactoring of Legacy JavaScript Code to ES6 Modules

The JavaScript language did not specify, until ECMAScript 6 (ES6), nativ...
research
03/11/2023

Fine-grained Visual Classification with High-temperature Refinement and Background Suppression

Fine-grained visual classification is a challenging task due to the high...
research
05/22/2020

Focus Longer to See Better:Recursively Refined Attention for Fine-Grained Image Classification

Deep Neural Network has shown great strides in the coarse-grained image ...
research
12/28/2022

Part-guided Relational Transformers for Fine-grained Visual Recognition

Fine-grained visual recognition is to classify objects with visually sim...
research
08/09/2022

Sports Video Analysis on Large-Scale Data

This paper investigates the modeling of automated machine description on...

Please sign up or login with your details

Forgot password? Click here to reset