ViTOL: Vision Transformer for Weakly Supervised Object Localization

04/14/2022
by   Saurav Gupta, et al.
13

Weakly supervised object localization (WSOL) aims at predicting object locations in an image using only image-level category labels. Common challenges that image classification models encounter when localizing objects are, (a) they tend to look at the most discriminative features in an image that confines the localization map to a very small region, (b) the localization maps are class agnostic, and the models highlight objects of multiple classes in the same image and, (c) the localization performance is affected by background noise. To alleviate the above challenges we introduce the following simple changes through our proposed method ViTOL. We leverage the vision-based transformer for self-attention and introduce a patch-based attention dropout layer (p-ADL) to increase the coverage of the localization map and a gradient attention rollout mechanism to generate class-dependent attention maps. We conduct extensive quantitative, qualitative and ablation experiments on the ImageNet-1K and CUB datasets. We achieve state-of-the-art MaxBoxAcc-V2 localization scores of 70.47 is available on https://github.com/Saurav-31/ViTOL

READ FULL TEXT

page 1

page 3

page 5

page 7

page 8

research
08/03/2022

Re-Attention Transformer for Weakly Supervised Object Localization

Weakly supervised object localization is a challenging task which aims t...
research
07/19/2020

A Generic Visualization Approach for Convolutional Neural Networks

Retrieval networks are essential for searching and indexing. Compared to...
research
03/18/2023

Spatial-Aware Token for Weakly Supervised Object Localization

Weakly supervised object localization (WSOL) is a challenging task aimin...
research
04/29/2021

MinMaxCAM: Improving object coverage for CAM-basedWeakly Supervised Object Localization

One of the most common problems of weakly supervised object localization...
research
07/19/2023

Generative Prompt Model for Weakly Supervised Object Localization

Weakly supervised object localization (WSOL) remains challenging when le...
research
08/03/2022

Statistical Attention Localization (SAL): Methodology and Application to Object Classification

A statistical attention localization (SAL) method is proposed to facilit...
research
06/09/2020

Rethinking Localization Map: Towards Accurate Object Perception with Self-Enhancement Maps

Recently, remarkable progress has been made in weakly supervised object ...

Please sign up or login with your details

Forgot password? Click here to reset