PatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and Visibility

08/19/2021
by   Jae Yong Lee, et al.
8

Recent learning-based multi-view stereo (MVS) methods show excellent performance with dense cameras and small depth ranges. However, non-learning based approaches still outperform for scenes with large depth ranges and sparser wide-baseline views, in part due to their PatchMatch optimization over pixelwise estimates of depth, normals, and visibility. In this paper, we propose an end-to-end trainable PatchMatch-based MVS approach that combines advantages of trainable costs and regularizations with pixelwise estimates. To overcome the challenge of the non-differentiable PatchMatch optimization that involves iterative sampling and hard decisions, we use reinforcement learning to minimize expected photometric cost and maximize likelihood of ground truth depth and normals. We incorporate normal estimation by using dilated patch kernels, and propose a recurrent cost regularization that applies beyond frontal plane-sweep algorithms to our pixelwise depth/normal estimates. We evaluate our method on widely used MVS benchmarks, ETH3D and Tanks and Temples (TnT), and compare to other state of the art learning based MVS models. On ETH3D, our method outperforms other recent learning-based approaches and performs comparably on advanced TnT.

READ FULL TEXT

page 1

page 3

page 4

page 6

page 8

research
10/14/2022

Deep PatchMatch MVS with Learned Patch Coplanarity, Geometric Consistency and Adaptive Pixel Sampling

Recent work in multi-view stereo (MVS) combines learnable photometric sc...
research
11/24/2019

Normal Assisted Stereo Depth Estimation

Accurate stereo depth estimation plays a critical role in various 3D tas...
research
11/29/2021

IB-MVS: An Iterative Algorithm for Deep Multi-View Stereo based on Binary Decisions

We present a novel deep-learning-based method for Multi-View Stereo. Our...
research
07/15/2020

PVSNet: Pixelwise Visibility-Aware Multi-View Stereo Network

Recently, learning-based multi-view stereo methods have achieved promisi...
research
08/04/2019

Adversarial View-Consistent Learning for Monocular Depth Estimation

This paper addresses the problem of Monocular Depth Estimation (MDE). Ex...
research
03/06/2020

DeLTra: Deep Light Transport for Projector-Camera Systems

In projector-camera systems, light transport models the propagation from...
research
08/17/2023

V-FUSE: Volumetric Depth Map Fusion with Long-Range Constraints

We introduce a learning-based depth map fusion framework that accepts a ...

Please sign up or login with your details

Forgot password? Click here to reset