Discrete Cosine Transform Network for Guided Depth Map Super-Resolution

04/14/2021
by   Zixiang Zhao, et al.
0

Guided depth super-resolution (GDSR) is a hot topic in multi-modal image processing. The goal is to use high-resolution (HR) RGB images to provide extra information on edges and object contours, so that low-resolution depth maps can be upsampled to HR ones. To solve the issues of RGB texture over-transferred, cross-modal feature extraction difficulty and unclear working mechanism of modules in existing methods, we propose an advanced Discrete Cosine Transform Network (DCTNet), which is composed of four components. Firstly, the paired RGB/depth images are input into the semi-coupled feature extraction module. The shared convolution kernels extract the cross-modal common features, and the private kernels extract their unique features, respectively. Then the RGB features are input into the edge attention mechanism to highlight the edges useful for upsampling. Subsequently, in the Discrete Cosine Transform (DCT) module, where DCT is employed to solve the optimization problem designed for image domain GDSR. The solution is then extended to implement the multi-channel RGB/depth features upsampling, which increases the rationality of DCTNet, and is more flexible and effective than conventional methods. The final depth prediction is output by the reconstruction module. Numerous qualitative and quantitative experiments demonstrate the effectiveness of our method, which can generate accurate and HR depth maps, surpassing state-of-the-art methods. Meanwhile, the rationality of modules is also proved by ablation experiments.

READ FULL TEXT

page 1

page 4

page 6

page 7

research
03/15/2023

Spherical Space Feature Decomposition for Guided Depth Map Super-Resolution

Guided depth map super-resolution (GDSR), as a hot topic in multi-modal ...
research
11/22/2019

PAG-Net: Progressive Attention Guided Depth Super-resolution Network

In this paper, we propose a novel method for the challenging problem of ...
research
04/04/2021

High-resolution Depth Maps Imaging via Attention-based Hierarchical Multi-modal Fusion

Depth map records distance between the viewpoint and objects in the scen...
research
02/28/2023

RGB-D Grasp Detection via Depth Guided Learning with Cross-modal Attention

Planar grasp detection is one of the most fundamental tasks to robotic m...
research
02/12/2020

Fast Generation of High Fidelity RGB-D Images by Deep-Learning with Adaptive Convolution

Using the raw data from consumer-level RGB-D cameras as input, we propos...
research
05/09/2017

Signal reconstruction via operator guiding

Signal reconstruction from a sample using an orthogonal projector onto a...
research
10/19/2020

Multi-Modal Super Resolution for Dense Microscopic Particle Size Estimation

Particle Size Analysis (PSA) is an important process carried out in a nu...

Please sign up or login with your details

Forgot password? Click here to reset