Probabilistic Graph Attention Network with Conditional Kernels for Pixel-Wise Prediction

01/08/2021
by   Dan Xu, et al.
17

Multi-scale representations deeply learned via convolutional neural networks have shown tremendous importance for various pixel-level prediction problems. In this paper we present a novel approach that advances the state of the art on pixel-level prediction in a fundamental aspect, i.e. structured multi-scale features learning and fusion. In contrast to previous works directly considering multi-scale feature maps obtained from the inner layers of a primary CNN architecture, and simply fusing the features with weighted averaging or concatenation, we propose a probabilistic graph attention network structure based on a novel Attention-Gated Conditional Random Fields (AG-CRFs) model for learning and fusing multi-scale representations in a principled manner. In order to further improve the learning capacity of the network structure, we propose to exploit feature dependant conditional kernels within the deep probabilistic framework. Extensive experiments are conducted on four publicly available datasets (i.e. BSDS500, NYUD-V2, KITTI, and Pascal-Context) and on three challenging pixel-wise prediction problems involving both discrete and continuous labels (i.e. monocular depth estimation, object contour prediction, and semantic segmentation). Quantitative and qualitative results demonstrate the effectiveness of the proposed latent AG-CRF model and the overall probabilistic graph attention network with feature conditional kernels for structured feature learning and pixel-wise prediction.

READ FULL TEXT

page 1

page 5

page 7

page 8

page 11

page 12

page 13

research
01/01/2018

Learning Deep Structured Multi-Scale Features using Attention-Gated CRFs for Contour Prediction

Recent works have shown that exploiting multi-scale representations deep...
research
03/29/2018

Structured Attention Guided Convolutional Neural Fields for Monocular Depth Estimation

Recent works have shown the benefit of integrating Conditional Random Fi...
research
09/08/2019

Squeeze-and-Attention Networks for Semantic Segmentation

Squeeze-and-excitation (SE) module enhances the representational power o...
research
03/05/2021

Variational Structured Attention Networks for Deep Visual Representation Learning

Convolutional neural networks have enabled major progress in addressing ...
research
10/30/2021

MFNet: Multi-class Few-shot Segmentation Network with Pixel-wise Metric Learning

In visual recognition tasks, few-shot learning requires the ability to l...
research
11/29/2017

Do Convolutional Neural Networks act as Compositional Nearest Neighbors?

We present a simple approach based on pixel-wise nearest neighbors to un...
research
05/15/2018

Automated Vision-based Bridge Component Extraction Using Multiscale Convolutional Neural Networks

Image data has a great potential of helping post-earthquake visual inspe...

Please sign up or login with your details

Forgot password? Click here to reset