Weakly Supervised Object Detection with Pointwise Mutual Information

01/26/2018
by   Rene Grzeszick, et al.
0

In this work a novel approach for weakly supervised object detection that incorporates pointwise mutual information is presented. A fully convolutional neural network architecture is applied in which the network learns one filter per object class. The resulting feature map indicates the location of objects in an image, yielding an intuitive representation of a class activation map. While traditionally such networks are learned by a softmax or binary logistic regression (sigmoid cross-entropy loss), a learning approach based on a cosine loss is introduced. A pointwise mutual information layer is incorporated in the network in order to project predictions and ground truth presence labels in a non-categorical embedding space. Thus, the cosine loss can be employed in this non-categorical representation. Besides integrating image level annotations, it is shown how to integrate point-wise annotations using a Spatial Pyramid Pooling layer. The approach is evaluated on the VOC2012 dataset for classification, point localization and weakly supervised bounding box localization. It is shown that the combination of pointwise mutual information and a cosine loss eases the learning process and thus improves the accuracy. The integration of coarse point-wise localizations further improves the results at minimal annotation costs.

READ FULL TEXT

page 4

page 7

research
11/25/2019

Rethinking Softmax with Cross-Entropy: Neural Network Classifier as Mutual Information Estimator

Mutual information is widely applied to learn latent representations of ...
research
06/19/2021

Neural Network Classifier as Mutual Information Evaluator

Cross-entropy loss with softmax output is a standard choice to train neu...
research
09/11/2019

WSOD^2: Learning Bottom-up and Top-down Objectness Distillation for Weakly-supervised Object Detection

We study on weakly-supervised object detection (WSOD) which plays a vita...
research
03/18/2020

In Defense of Graph Inference Algorithms for Weakly Supervised Object Localization

Weakly Supervised Object Localization (WSOL) methods have become increas...
research
03/03/2015

Weakly Supervised Object Localization with Multi-fold Multiple Instance Learning

Object category localization is a challenging problem in computer vision...
research
05/09/2019

Learning Interpretable Features via Adversarially Robust Optimization

Neural networks are proven to be remarkably successful for classificatio...
research
04/30/2019

Investigation of Initialization Strategies for the Multiple Instance Adaptive Cosine Estimator

Sensors which use electromagnetic induction (EMI) to excite a response i...

Please sign up or login with your details

Forgot password? Click here to reset