STS: Surround-view Temporal Stereo for Multi-view 3D Detection

08/22/2022
by   Zengran Wang, et al.
0

Learning accurate depth is essential to multi-view 3D object detection. Recent approaches mainly learn depth from monocular images, which confront inherent difficulties due to the ill-posed nature of monocular depth learning. Instead of using a sole monocular depth method, in this work, we propose a novel Surround-view Temporal Stereo (STS) technique that leverages the geometry correspondence between frames across time to facilitate accurate depth learning. Specifically, we regard the field of views from all cameras around the ego vehicle as a unified view, namely surroundview, and conduct temporal stereo matching on it. The resulting geometrical correspondence between different frames from STS is utilized and combined with the monocular depth to yield final depth prediction. Comprehensive experiments on nuScenes show that STS greatly boosts 3D detection ability, notably for medium and long distance objects. On BEVDepth with ResNet-50 backbone, STS improves mAP and NDS by 2.6 and 1.4 larger backbone and a larger image resolution, demonstrating its effectiveness

READ FULL TEXT

page 2

page 4

page 7

research
07/26/2022

MV-FCOS3D++: Multi-View Camera-Only 4D Object Detection with Pretrained Monocular Backbones

In this technical report, we present our solution, dubbed MV-FCOS3D++, f...
research
08/17/2019

Mono-SF: Multi-View Geometry Meets Single-View Depth for Monocular Scene Flow Estimation of Dynamic Traffic Scenes

Existing 3D scene flow estimation methods provide the 3D geometry and 3D...
research
07/26/2022

Monocular 3D Object Detection with Depth from Motion

Perceiving 3D objects from monocular inputs is crucial for robotic syste...
research
12/17/2019

Single-Stage Monocular 3D Object Detection with Virtual Cameras

While expensive LiDAR and stereo camera rigs have enabled the developmen...
research
03/07/2018

Single View Stereo Matching

Previous monocular depth estimation methods take a single view and direc...
research
03/22/2018

Prioritized Multi-View Stereo Depth Map Generation Using Confidence Prediction

In this work, we propose a novel approach to prioritize the depth map co...
research
10/05/2022

Time Will Tell: New Outlooks and A Baseline for Temporal Multi-View 3D Object Detection

While recent camera-only 3D detection methods leverage multiple timestep...

Please sign up or login with your details

Forgot password? Click here to reset