Scene Recognition with Prototype-agnostic Scene Layout

09/07/2019
by   Gongwei Chen, et al.
0

Abstract--- Exploiting the spatial structure in scene images is a key research direction for scene recognition. Due to the large intra-class structural diversity, building and modeling flexible structural layout to adapt various image characteristics is a challenge. Existing structural modeling methods in scene recognition either focus on predefined grids or rely on learned prototypes, which all have limited representative ability. In this paper, we propose Prototype-agnostic Scene Layout (PaSL) construction method to build the spatial structure for each image without conforming to any prototype. Our PaSL can flexibly capture the diverse spatial characteristic of scene images and have considerable generalization capability. Given a PaSL, we build Layout Graph Network (LGN) where regions in PaSL are defined as nodes and two kinds of independent relations between regions are encoded as edges. The LGN aims to incorporate two topological structures (formed in spatial and semantic similarity dimensions) into image representations through graph convolution. Extensive experiments show that our approach achieves state-of-the-art results on widely recognized MIT67 and SUN397 datasets without multi-model or multi-scale fusion. Moreover, we also conduct the experiments on one of the largest scale datasets, Places365. The results demonstrate the proposed method can be well generalized and obtains competitive performance.

READ FULL TEXT

page 1

page 4

page 8

research
09/02/2019

Relationship-Aware Spatial Perception Fusion for Realistic Scene Layout Generation

The significant progress on Generative Adversarial Networks (GANs) have ...
research
05/05/2022

Scene Graph Expansion for Semantics-Guided Image Outpainting

In this paper, we address the task of semantics-guided image outpainting...
research
07/10/2018

Deep Structured Generative Models

Deep generative models have shown promising results in generating realis...
research
08/12/2022

Layout-Bridging Text-to-Image Synthesis

The crux of text-to-image synthesis stems from the difficulty of preserv...
research
10/09/2021

SGMNet: Scene Graph Matching Network for Few-Shot Remote Sensing Scene Classification

Few-Shot Remote Sensing Scene Classification (FSRSSC) is an important ta...
research
06/18/2015

A Spatial Layout and Scale Invariant Feature Representation for Indoor Scene Classification

Unlike standard object classification, where the image to be classified ...
research
08/19/2019

Seq-SG2SL: Inferring Semantic Layout from Scene Graph Through Sequence to Sequence Learning

Generating semantic layout from scene graph is a crucial intermediate ta...

Please sign up or login with your details

Forgot password? Click here to reset