Self-Supervised Road Layout Parsing with Graph Auto-Encoding

03/21/2022
by   Chenyang Lu, et al.
0

Aiming for higher-level scene understanding, this work presents a neural network approach that takes a road-layout map in bird's eye view as input, and predicts a human-interpretable graph that represents the road's topological layout. Our approach elevates the understanding of road layouts from pixel level to the level of graphs. To achieve this goal, an image-graph-image auto-encoder is utilized. The network is designed to learn to regress the graph representation at its auto-encoder bottleneck. This learning is self-supervised by an image reconstruction loss, without needing any external manual annotations. We create a synthetic dataset containing common road layout patterns and use it for training of the auto-encoder in addition to the real-world Argoverse dataset. By using this additional synthetic dataset, which conceptually captures human knowledge of road layouts and makes this available to the network for training, we are able to stabilize and further improve the performance of topological road layout understanding on the real-world Argoverse dataset. The evaluation shows that our approach exhibits comparable performance to a strong fully-supervised baseline.

READ FULL TEXT

page 1

page 4

page 5

page 7

page 8

page 9

research
12/10/2020

Image-Graph-Image Translation via Auto-Encoding

This work presents the first convolutional neural network that learns an...
research
03/30/2022

Self-supervised 360^∘ Room Layout Estimation

We present the first self-supervised method to train panoramic room layo...
research
11/22/2021

Efficient Non-Compression Auto-Encoder for Driving Noise-based Road Surface Anomaly Detection

Wet weather makes water film over the road and that film causes lower fr...
research
02/19/2020

MonoLayout: Amodal scene layout from a single image

In this paper, we address the novel, highly challenging problem of estim...
research
12/15/2021

Gaze Estimation with Eye Region Segmentation and Self-Supervised Multistream Learning

We present a novel multistream network that learns robust eye representa...
research
12/06/2018

Auto-Encoding Scene Graphs for Image Captioning

We propose Scene Graph Auto-Encoder (SGAE) that incorporates the languag...
research
12/06/2018

Auto-Encoding Graphical Inductive Bias for Descriptive Image Captioning

We propose Scene Graph Auto-Encoder (SGAE) that incorporates the languag...

Please sign up or login with your details

Forgot password? Click here to reset