ES-MVSNet: Efficient Framework for End-to-end Self-supervised Multi-View Stereo

08/04/2023
by   Qiang Zhou, et al.
0

Compared to the multi-stage self-supervised multi-view stereo (MVS) method, the end-to-end (E2E) approach has received more attention due to its concise and efficient training pipeline. Recent E2E self-supervised MVS approaches have integrated third-party models (such as optical flow models, semantic segmentation models, NeRF models, etc.) to provide additional consistency constraints, which grows GPU memory consumption and complicates the model's structure and training pipeline. In this work, we propose an efficient framework for end-to-end self-supervised MVS, dubbed ES-MVSNet. To alleviate the high memory consumption of current E2E self-supervised MVS frameworks, we present a memory-efficient architecture that reduces memory usage by 43 without compromising model performance. Furthermore, with the novel design of asymmetric view selection policy and region-aware depth consistency, we achieve state-of-the-art performance among E2E self-supervised MVS methods, without relying on third-party models for additional consistency signals. Extensive experiments on DTU and Tanks Temples benchmarks demonstrate that the proposed ES-MVSNet approach achieves state-of-the-art performance among E2E self-supervised MVS methods and competitive performance to many supervised and multi-stage self-supervised methods.

READ FULL TEXT

page 2

page 5

page 7

research
08/30/2021

Digging into Uncertainty in Self-supervised Multi-view Stereo

Self-supervised Multi-view stereo (MVS) with a pretext task of image rec...
research
04/12/2021

Self-supervised Multi-view Stereo via Effective Co-Segmentation and Data-Augmentation

Recent studies have witnessed that self-supervised methods based on view...
research
07/02/2023

End-to-End Out-of-distribution Detection with Self-supervised Sampling

Out-of-distribution (OOD) detection empowers the model trained on the cl...
research
03/08/2022

End-to-end Multiple Instance Learning with Gradient Accumulation

Being able to learn on weakly labeled data, and provide interpretability...
research
02/07/2023

Scaling Self-Supervised End-to-End Driving with Multi-View Attention Learning

On end-to-end driving, a large amount of expert driving demonstrations i...
research
05/29/2023

View-to-Label: Multi-View Consistency for Self-Supervised 3D Object Detection

For autonomous vehicles, driving safely is highly dependent on the capab...
research
01/13/2021

Self-Supervised Vessel Enhancement Using Flow-Based Consistencies

Vessel segmenting is an essential task in many clinical applications. Al...

Please sign up or login with your details

Forgot password? Click here to reset