OCNet: Object Context Network for Scene Parsing

09/04/2018
by   Yuhui Yuan, et al.
0

Context is essential for various computer vision tasks. The state-of-the-art scene parsing methods have exploited the effectiveness of the context defined over image-level. Such context carries the mixture of objects belonging to different categories. According to that the label of each pixel P is defined as the category of the object it belongs to, we propose the pixel-wise Object Context that consists of the objects belonging to the same category with pixel P. The representation of pixel P's object context is the aggregation of all the features that belong to the pixels sharing the same category with P. Since the ground truth objects that the pixel P belonging to is unavailable, we employ the self-attention method to approximate the objects by learning a pixel-wise similarity map. We further propose the Pyramid Object Context and Atrous Spatial Pyramid Object Context to capture context of multiple scales. Based on the object context, we introduce the OCNet and show that OCNet achieves state-of-the-art performance on both Cityscapes benchmark and ADE20K benchmark. The code of OCNet will be made available at https://github.com/PkuRainBow/OCNet.

READ FULL TEXT
research
12/04/2016

Pyramid Scene Parsing Network

Scene parsing is challenging for unrestricted open vocabulary and divers...
research
03/14/2023

Co-Salient Object Detection with Co-Representation Purification

Co-salient object detection (Co-SOD) aims at discovering the common obje...
research
07/12/2022

Tracking Objects as Pixel-wise Distributions

Multi-object tracking (MOT) requires detecting and associating objects t...
research
03/29/2023

DPF: Learning Dense Prediction Fields with Weak Supervision

Nowadays, many visual scene understanding problems are addressed by dens...
research
09/12/2019

Detecting Robotic Affordances on Novel Objects with Regional Attention and Attributes

This paper presents a framework for predicting affordances of object par...
research
02/01/2019

Lift-the-Flap: Context Reasoning Using Object-Centered Graphs

Children benefit from lift-the-flap books by taking on an active role in...
research
03/07/2022

GlideNet: Global, Local and Intrinsic based Dense Embedding NETwork for Multi-category Attributes Prediction

Attaching attributes (such as color, shape, state, action) to object cat...

Please sign up or login with your details

Forgot password? Click here to reset