Spatio-Temporal Video Representation Learning for AI Based Video Playback Style Prediction

10/03/2021
by   Gaurav Ramola, et al.
1

Ever-increasing smartphone-generated video content demands intelligent techniques to edit and enhance videos on power-constrained devices. Most of the best performing algorithms for video understanding tasks like action recognition, localization, etc., rely heavily on rich spatio-temporal representations to make accurate predictions. For effective learning of the spatio-temporal representation, it is crucial to understand the underlying object motion patterns present in the video. In this paper, we propose a novel approach for understanding object motions via motion type classification. The proposed motion type classifier predicts a motion type for the video based on the trajectories of the objects present. Our classifier assigns a motion type for the given video from the following five primitive motion classes: linear, projectile, oscillatory, local and random. We demonstrate that the representations learned from the motion type classification generalizes well for the challenging downstream task of video retrieval. Further, we proposed a recommendation system for video playback style based on the motion type classifier predictions.

READ FULL TEXT

page 1

page 3

page 6

page 8

research
05/10/2019

Spatio-temporal Video Re-localization by Warp LSTM

The need for efficiently finding the video content a user wants is incre...
research
02/14/2021

Learning Self-Similarity in Space and Time as Generalized Motion for Action Recognition

Spatio-temporal convolution often fails to learn motion dynamics in vide...
research
09/29/2020

Robust Detection of Objects under Periodic Motion with Gaussian Process Filtering

Object Detection (OD) is an important task in Computer Vision with many ...
research
05/01/2014

Retrieval in Long Surveillance Videos using User Described Motion and Object Attributes

We present a content-based retrieval method for long surveillance videos...
research
05/13/2023

Lightweight Delivery Detection on Doorbell Cameras

Despite recent advances in video-based action recognition and robust spa...
research
01/27/2016

Deep Learning Driven Visual Path Prediction from a Single Image

Capabilities of inference and prediction are significant components of v...
research
06/18/2020

Learning non-rigid surface reconstruction from spatio-temporal image patches

We present a method to reconstruct a dense spatio-temporal depth map of ...

Please sign up or login with your details

Forgot password? Click here to reset