Stereo Matching in Time: 100+ FPS Video Stereo Matching for Extended Reality

09/08/2023
by   Ziang Cheng, et al.
0

Real-time Stereo Matching is a cornerstone algorithm for many Extended Reality (XR) applications, such as indoor 3D understanding, video pass-through, and mixed-reality games. Despite significant advancements in deep stereo methods, achieving real-time depth inference with high accuracy on a low-power device remains a major challenge. One of the major difficulties is the lack of high-quality indoor video stereo training datasets captured by head-mounted VR/AR glasses. To address this issue, we introduce a novel video stereo synthetic dataset that comprises photorealistic renderings of various indoor scenes and realistic camera motion captured by a 6-DoF moving VR/AR head-mounted display (HMD). This facilitates the evaluation of existing approaches and promotes further research on indoor augmented reality scenarios. Our newly proposed dataset enables us to develop a novel framework for continuous video-rate stereo matching. As another contribution, our dataset enables us to proposed a new video-based stereo matching approach tailored for XR applications, which achieves real-time inference at an impressive 134fps on a standard desktop computer, or 30fps on a battery-powered HMD. Our key insight is that disparity and contextual information are highly correlated and redundant between consecutive stereo frames. By unrolling an iterative cost aggregation in time (i.e. in the temporal dimension), we are able to distribute and reuse the aggregated features over time. This approach leads to a substantial reduction in computation without sacrificing accuracy. We conducted extensive evaluations and comparisons and demonstrated that our method achieves superior performance compared to the current state-of-the-art, making it a strong contender for real-time stereo matching in VR/AR applications.

READ FULL TEXT

page 1

page 7

research
01/01/2021

Bilateral Grid Learning for Stereo Matching Network

The real-time performance of the stereo matching network is important fo...
research
08/03/2017

Real-time Geometry-Aware Augmented Reality in Minimally Invasive Surgery

The potential of Augmented Reality (AR) technology to assist minimally i...
research
04/24/2023

Auto-CARD: Efficient and Robust Codec Avatar Driving for Real-time Mobile Telepresence

Real-time and robust photorealistic avatars for telepresence in AR/VR ha...
research
08/14/2020

MatryODShka: Real-time 6DoF Video View Synthesis using Multi-Sphere Images

We introduce a method to convert stereo 360 (omnidirectional stereo) ima...
research
02/17/2019

Exploring Stereovision-Based 3-D Scene Reconstruction for Augmented Reality

Three-dimensional (3-D) scene reconstruction is one of the key technique...
research
06/22/2022

A High Resolution Multi-exposure Stereoscopic Image Video Database of Natural Scenes

Immersive displays such as VR headsets, AR glasses, Multiview displays, ...
research
09/22/2017

Virtual Blood Vessels in Complex Background using Stereo X-ray Images

We propose a fully automatic system to reconstruct and visualize 3D bloo...

Please sign up or login with your details

Forgot password? Click here to reset