Masked GANs for Unsupervised Depth and Pose Prediction with Scale Consistency

04/09/2020
by   Chaoqiang Zhao, et al.
0

Previous works have shown that adversarial learning can be used for unsupervised monocular depth and visual odometry (VO) estimation. However, the performance of pose and depth networks is limited by occlusions and visual field changes. Because of the incomplete correspondence of visual information between frames caused by motion, target images cannot be synthesized completely from source images via view reconstruction and bilinear interpolation. The reconstruction loss based on the difference between synthesized and real target images will be affected by the incomplete reconstruction. Besides, the data distribution of unreconstructed regions will be learned and help the discriminator distinguish between real and fake images, thereby causing the case that the generator may fail to compete with the discriminator. Therefore, a MaskNet is designed in this paper to predict these regions and reduce their impacts on the reconstruction loss and adversarial loss. The impact of unreconstructed regions on discriminator is tackled by proposing a boolean mask scheme, as shown in Fig. 1. Furthermore, we consider the scale consistency of our pose network by utilizing a new scale-consistency loss, therefore our pose network is capable of providing the full camera trajectory over the long monocular sequence. Extensive experiments on KITTI dataset show that each component proposed in this paper contributes to the performance, and both of our depth and trajectory prediction achieve competitive performance.

READ FULL TEXT

page 1

page 2

page 4

page 6

page 7

page 8

research
09/20/2017

UnDeepVO: Monocular Visual Odometry through Unsupervised Deep Learning

We propose a novel monocular visual odometry (VO) system called UnDeepVO...
research
08/28/2019

Unsupervised Scale-consistent Depth and Ego-motion Learning from Monocular Video

Recent work has shown that CNN-based depth and ego-motion estimators can...
research
09/16/2018

GANVO: Unsupervised Deep Monocular Visual Odometry and Depth Estimation with Generative Adversarial Networks

In the last decade, supervised deep learning approaches have been extens...
research
03/03/2020

DiPE: Deeper into Photometric Errors for Unsupervised Learning of Depth and Ego-motion from Monocular Videos

Unsupervised learning of depth and ego-motion from unlabelled monocular ...
research
03/09/2019

Sparse Representations for Object and Ego-motion Estimation in Dynamic Scenes

Dynamic scenes that contain both object motion and egomotion are a chall...
research
03/01/2021

Unsupervised Depth and Ego-motion Estimation for Monocular Thermal Video using Multi-spectral Consistency Loss

Most of the deep-learning based depth and ego-motion networks have been ...
research
11/02/2020

Unsupervised Monocular Depth Learning with Integrated Intrinsics and Spatio-Temporal Constraints

Monocular depth inference has gained tremendous attention from researche...

Please sign up or login with your details

Forgot password? Click here to reset