DeepRelativeFusion: Dense Monocular SLAM using Single-Image Relative Depth Prediction

06/07/2020
by   Shing Yan Loo, et al.
0

Traditional monocular visual simultaneous localization and mapping (SLAM) algorithms have been extensively studied and proven to reliably recover a sparse structure and camera motion. Nevertheless, the sparse structure is still insufficient for scene interaction, e.g., visual navigation and augmented reality applications. To densify the scene reconstruction, the use of single-image absolute depth prediction from convolutional neural networks (CNNs) for filling in the missing structure has been proposed. However, the prediction accuracy tends to not generalize well on scenes that are different from the training datasets. In this paper, we propose a dense monocular SLAM system, named DeepRelativeFusion, that is capable to recover a globally consistent 3D structure. To this end, we use a visual SLAM algorithm to reliably recover the camera poses and semi-dense depth maps of the keyframes, and then combine the keyframe pose-graph with the densified keyframe depth maps to reconstruct the scene. To perform the densification, we introduce two incremental improvements upon the energy minimization framework proposed by DeepFusion: (1) an additional image gradient term in the cost function, and (2) the use of single-image relative depth prediction. Despite the absence of absolute scale and depth range, the relative depth maps can be corrected using their respective semi-dense depth maps from the SLAM algorithm. We show that the corrected relative depth maps are sufficiently accurate to be used as priors for the densification. To demonstrate the generalizability of relative depth prediction, we illustrate qualitatively the dense reconstruction on two outdoor sequences. Our system also outperforms the state-of-the-art dense SLAM systems quantitatively in dense reconstruction accuracy by a large margin.

READ FULL TEXT

page 1

page 6

page 7

research
04/11/2017

CNN-SLAM: Real-time dense monocular SLAM with learned depth prediction

Given the recent advances in depth prediction from Convolutional Neural ...
research
08/17/2021

A Hybrid Sparse-Dense Monocular SLAM System for Autonomous Driving

In this paper, we present a system for incrementally reconstructing a de...
research
07/30/2018

Geo-Supervised Visual Depth Prediction

We propose using global orientation from inertial measurements, and the ...
research
08/20/2021

Deep Virtual Markers for Articulated 3D Shapes

We propose deep virtual markers, a framework for estimating dense and ac...
research
03/01/2017

Augmented Reality for Depth Cues in Monocular Minimally Invasive Surgery

One of the major challenges in Minimally Invasive Surgery (MIS) such as ...
research
01/17/2019

Towards Building the Semantic Map from a Monocular Camera with a Multi-task Network

In many robotic applications, especially for the autonomous driving, und...
research
09/21/2017

Sparse-to-Dense: Depth Prediction from Sparse Depth Samples and a Single Image

We consider the problem of dense depth prediction from a sparse set of d...

Please sign up or login with your details

Forgot password? Click here to reset