Deep Homography Estimation for Dynamic Scenes

04/05/2020
by   Hoang Le, et al.
4

Homography estimation is an important step in many computer vision problems. Recently, deep neural network methods have shown to be favorable for this problem when compared to traditional methods. However, these new methods do not consider dynamic content in input images. They train neural networks with only image pairs that can be perfectly aligned using homographies. This paper investigates and discusses how to design and train a deep neural network that handles dynamic scenes. We first collect a large video dataset with dynamic content. We then develop a multi-scale neural network and show that when properly trained using our new dataset, this neural network can already handle dynamic scenes to some extent. To estimate a homography of a dynamic scene in a more principled way, we need to identify the dynamic content. Since dynamic content detection and homography estimation are two tightly coupled tasks, we follow the multi-task learning principles and augment our multi-scale network such that it jointly estimates the dynamics masks and homographies. Our experiments show that our method can robustly estimate homography for challenging scenarios with dynamic scenes, blur artifacts, or lack of textures.

READ FULL TEXT

page 1

page 3

page 6

page 8

research
12/07/2016

Deep Multi-scale Convolutional Neural Network for Dynamic Scene Deblurring

Non-uniform blind deblurring for general dynamic scenes is a challenging...
research
08/04/2022

Multi-scale Sampling and Aggregation Network For High Dynamic Range Imaging

High dynamic range (HDR) imaging is a fundamental problem in image proce...
research
07/06/2022

Learning Regularized Multi-Scale Feature Flow for High Dynamic Range Imaging

Reconstructing ghosting-free high dynamic range (HDR) images of dynamic ...
research
11/27/2021

Video Content Classification using Deep Learning

Video content classification is an important research content in compute...
research
04/25/2019

Holistic Large Scale Video Understanding

Action recognition has been advanced in recent years by benchmarks with ...
research
05/01/2023

Joint tone mapping and denoising of thermal infrared images via multi-scale Retinex and multi-task learning

Cameras digitize real-world scenes as pixel intensity values with a limi...
research
11/30/2022

Two-branch Multi-scale Deep Neural Network for Generalized Document Recapture Attack Detection

The image recapture attack is an effective image manipulation method to ...

Please sign up or login with your details

Forgot password? Click here to reset