Horizontal-to-Vertical Video Conversion

01/11/2021
by   Tun Zhu, et al.
12

Alongside the prevalence of mobile videos, the general public leans towards consuming vertical videos on hand-held devices. To revitalize the exposure of horizontal contents, we hereby set forth the exploration of automated horizontal-to-vertical (abbreviated as H2V) video conversion with our proposed H2V framework, accompanied by an accurately annotated H2V-142K dataset. Concretely, H2V framework integrates video shot boundary detection, subject selection and multi-object tracking to facilitate the subject-preserving conversion, wherein the key is subject selection. To achieve so, we propose a Rank-SS module that detects human objects, then selects the subject-to-preserve via exploiting location, appearance, and salient cues. Afterward, the framework automatically crops the video around the subject to produce vertical contents from horizontal sources. To build and evaluate our H2V framework, H2V-142K dataset is densely annotated with subject bounding boxes for 125 videos with 132K frames and 9,500 video covers, upon which we demonstrate superior subject selection performance comparing to traditional salient approaches, and exhibit promising horizontal-to-vertical conversion performance overall. By publicizing this dataset as well as our approach, we wish to pave the way for more valuable endeavors on the horizontal-to-vertical video conversion task.

READ FULL TEXT

page 1

page 2

page 4

page 5

page 6

page 10

page 11

research
05/24/2021

SHD360: A Benchmark Dataset for Salient Human Detection in 360° Videos

Salient human detection (SHD) in dynamic 360 immersive videos is of grea...
research
04/05/2022

Learning Video Salient Object Detection Progressively from Unlabeled Videos

Recent deep learning-based video salient object detection (VSOD) has ach...
research
02/02/2017

YouTube-BoundingBoxes: A Large High-Precision Human-Annotated Data Set for Object Detection in Video

We introduce a new large-scale data set of video URLs with densely-sampl...
research
07/24/2021

ASOD60K: Audio-Induced Salient Object Detection in Panoramic Videos

Exploring to what humans pay attention in dynamic panoramic scenes is us...
research
11/26/2018

Foreground Clustering for Joint Segmentation and Localization in Videos and Images

This paper presents a novel framework in which video/image segmentation ...
research
04/13/2016

Deep3D: Fully Automatic 2D-to-3D Video Conversion with Deep Convolutional Neural Networks

As 3D movie viewing becomes mainstream and Virtual Reality (VR) market e...
research
07/24/2019

StableNet: Semi-Online, Multi-Scale Deep Video Stabilization

Video stabilization algorithms are of greater importance nowadays with t...

Please sign up or login with your details

Forgot password? Click here to reset