MoNet: Moments Embedding Network

02/20/2018
by   Mengran Gou, et al.
0

Bilinear pooling has been recently proposed as a feature encoding layer, which can be used after the convolutional layers of a deep network, to improve performance in multiple vision tasks. Instead of conventional global average pooling or fully connected layer, bilinear pooling gathers 2nd order information in a translation invariant fashion. However, a serious drawback of this family of pooling layers is their dimensionality explosion. Approximate pooling methods with compact property have been explored towards resolving this weakness. Additionally, recent results have shown that significant performance gains can be achieved by using matrix normalization to regularize unstable higher order information. However, combining compact pooling with matrix normalization has not been explored until now. In this paper, we unify the bilinear pooling layer and the global Gaussian embedding layer through the empirical moment matrix. In addition, with a proposed novel sub-matrix square-root layer, one can normalize the output of the convolution layer directly and mitigate the dimensionality problem with off-the-shelf compact pooling methods. Our experiments on three widely used fine-grained classification datasets illustrate that our proposed architecture MoNet can achieve similar or better performance than G2DeNet . When combined with compact pooling technique, it obtains comparable performance with the encoded feature of 96

READ FULL TEXT
POST COMMENT

Comments

There are no comments yet.

Authors

page 4

07/26/2018

Hierarchical Bilinear Pooling for Fine-Grained Visual Recognition

Fine-grained visual recognition is challenging because it highly relies ...
11/16/2016

Low-rank Bilinear Pooling for Fine-Grained Classification

Pooling second-order local feature statistics to form a high-dimensional...
06/05/2019

Compact Approximation for Polynomial of Covariance Feature

Covariance pooling is a feature pooling method with good classification ...
12/04/2017

Towards Faster Training of Global Covariance Pooling Networks by Iterative Matrix Square Root Normalization

Global covariance pooling in Convolutional neural neworks has achieved i...
07/21/2017

Improved Bilinear Pooling with CNNs

Bilinear pooling of Convolutional Neural Network (CNN) features [22, 23]...
12/05/2018

Local Temporal Bilinear Pooling for Fine-grained Action Parsing

Fine-grained temporal action parsing is important in many applications, ...
12/16/2015

Multiregion Bilinear Convolutional Neural Networks for Person Re-Identification

In this work we propose a new architecture for person re-identification....
This week in AI

Get the week's most popular data science and artificial intelligence research sent straight to your inbox every Saturday.