Future Frame Prediction of a Video Sequence

08/31/2020
by   Jasmeen Kaur, et al.
0

Predicting future frames of a video sequence has been a problem of high interest in the field of Computer Vision as it caters to a multitude of applications. The ability to predict, anticipate and reason about future events is the essence of intelligence and one of the main goals of decision-making systems such as human-machine interaction, robot navigation and autonomous driving. However, the challenge lies in the ambiguous nature of the problem as there may be multiple future sequences possible for the same input video shot. A naively designed model averages multiple possible futures into a single blurry prediction. Recently, two distinct approaches have attempted to address this problem as: (a) use of latent variable models that represent underlying stochasticity and (b) adversarially trained models that aim to produce sharper images. A latent variable model often struggles to produce realistic results, while an adversarially trained model underutilizes latent variables and thus fails to produce diverse predictions. These methods have revealed complementary strengths and weaknesses. Combining the two approaches produces predictions that appear more realistic and better cover the range of plausible futures. This forms the basis and objective of study in this project work. In this paper, we proposed a novel multi-scale architecture combining both approaches. We validate our proposed model through a series of experiments and empirical evaluations on Moving MNIST, UCF101, and Penn Action datasets. Our method outperforms the results obtained using the baseline methods.

READ FULL TEXT
research
04/04/2018

Stochastic Adversarial Video Prediction

Being able to predict what may happen in the future requires an in-depth...
research
04/27/2019

Improved Conditional VRNNs for Video Prediction

Predicting future frames for a video sequence is a challenging generativ...
research
12/15/2019

Brain-Inspired Inference on Missing Video Sequence

In this paper, we propose a novel end-to-end architecture that could gen...
research
06/25/2016

An Uncertain Future: Forecasting from Static Images using Variational Autoencoders

In a given scene, humans can often easily predict a set of immediate fut...
research
06/05/2023

Introduction to Latent Variable Energy-Based Models: A Path Towards Autonomous Machine Intelligence

Current automated systems have crucial limitations that need to be addre...
research
10/21/2019

Learning to Make Generalizable and Diverse Predictions for Retrosynthesis

We propose a new model for making generalizable and diverse retrosynthet...
research
06/20/2018

Accurate and Diverse Sampling of Sequences based on a "Best of Many" Sample Objective

For autonomous agents to successfully operate in the real world, anticip...

Please sign up or login with your details

Forgot password? Click here to reset