Structure-Aware Residual Pyramid Network for Monocular Depth Estimation

07/13/2019
by   Xiaotian Chen, et al.
2

Monocular depth estimation is an essential task for scene understanding. The underlying structure of objects and stuff in a complex scene is critical to recovering accurate and visually-pleasing depth maps. Global structure conveys scene layouts, while local structure reflects shape details. Recently developed approaches based on convolutional neural networks (CNNs) significantly improve the performance of depth estimation. However, few of them take into account multi-scale structures in complex scenes. In this paper, we propose a Structure-Aware Residual Pyramid Network (SARPN) to exploit multi-scale structures for accurate depth prediction. We propose a Residual Pyramid Decoder (RPD) which expresses global scene structure in upper levels to represent layouts, and local structure in lower levels to present shape details. At each level, we propose Residual Refinement Modules (RRM) that predict residual maps to progressively add finer structures on the coarser structure predicted at the upper level. In order to fully exploit multi-scale image features, an Adaptive Dense Feature Fusion (ADFF) module, which adaptively fuses effective features from all scales for inferring structures of each scale, is introduced. Experiment results on the challenging NYU-Depth v2 dataset demonstrate that our proposed approach achieves state-of-the-art performance in both qualitative and quantitative evaluation. The code is available at https://github.com/Xt-Chen/SARPN.

READ FULL TEXT

page 1

page 3

page 4

page 5

page 6

research
01/19/2022

Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth

Depth estimation from a single image is an important task that can be ap...
research
04/05/2022

Pyramid Frequency Network with Spatial Attention Residual Refinement Module for Monocular Depth Estimation

Deep-learning-based approaches to depth estimation are rapidly advancing...
research
03/09/2022

Monocular Depth Distribution Alignment with Low Computation

The performance of monocular depth estimation generally depends on the a...
research
09/28/2018

Depth Reconstruction of Translucent Objects from a Single Time-of-Flight Camera using Deep Residual Networks

We propose a novel approach to recovering the translucent objects from a...
research
05/16/2020

Deep feature fusion for self-supervised monocular depth prediction

Recent advances in end-to-end unsupervised learning has significantly im...
research
02/26/2021

Boundary-induced and scene-aggregated network for monocular depth prediction

Monocular depth prediction is an important task in scene understanding. ...
research
03/15/2023

Skinned Motion Retargeting with Residual Perception of Motion Semantics Geometry

A good motion retargeting cannot be reached without reasonable considera...

Please sign up or login with your details

Forgot password? Click here to reset