Improving the generalization of network based relative pose regression: dimension reduction as a regularizer

10/24/2020
by   Xiaqing Ding, et al.
0

Visual localization occupies an important position in many areas such as Augmented Reality, robotics and 3D reconstruction. The state-of-the-art visual localization methods perform pose estimation using geometry based solver within the RANSAC framework. However, these methods require accurate pixel-level matching at high image resolution, which is hard to satisfy under significant changes from appearance, dynamics or perspective of view. End-to-end learning based regression networks provide a solution to circumvent the requirement for precise pixel-level correspondences, but demonstrate poor performance towards cross-scene generalization. In this paper, we explicitly add a learnable matching layer within the network to isolate the pose regression solver from the absolute image feature values, and apply dimension regularization on both the correlation feature channel and the image scale to further improve performance towards generalization and large viewpoint change. We implement this dimension regularization strategy within a two-layer pyramid based framework to regress the localization results from coarse to fine. In addition, the depth information is fused for absolute translational scale recovery. Through experiments on real world RGBD datasets we validate the effectiveness of our design in terms of improving both generalization performance and robustness towards viewpoint change, and also show the potential of regression based visual localization networks towards challenging occasions that are difficult for geometry based visual localization methods.

READ FULL TEXT

page 1

page 5

research
12/09/2017

SPP-Net: Deep Absolute Pose Regression with Synthetic Views

Image based localization is one of the important problems in computer vi...
research
09/14/2023

EP2P-Loc: End-to-End 3D Point to 2D Pixel Localization for Large-Scale Visual Localization

Visual localization is the task of estimating a 6-DoF camera pose of a q...
research
09/04/2018

Leveraging Deep Visual Descriptors for Hierarchical Efficient Localization

Many robotics applications require precise pose estimates despite operat...
research
07/01/2021

Deep auxiliary learning for visual localization using colorization task

Visual localization is one of the most important components for robotics...
research
03/28/2023

Cross-View Visual Geo-Localization for Outdoor Augmented Reality

Precise estimation of global orientation and location is critical to ens...
research
10/11/2022

DeepMLE: A Robust Deep Maximum Likelihood Estimator for Two-view Structure from Motion

Two-view structure from motion (SfM) is the cornerstone of 3D reconstruc...
research
04/17/2023

DeepSim-Nets: Deep Similarity Networks for Stereo Image Matching

We present three multi-scale similarity learning architectures, or DeepS...

Please sign up or login with your details

Forgot password? Click here to reset