WinDB: HMD-free and Distortion-free Panoptic Video Fixation Learning

05/23/2023
by   Guotao Wang, et al.
0

To date, the widely-adopted way to perform fixation collection in panoptic video is based on a head-mounted display (HMD), where participants' fixations are collected while wearing an HMD to explore the given panoptic scene freely. However, this widely-used data collection method is insufficient for training deep models to accurately predict which regions in a given panoptic are most important when it contains intermittent salient events. The main reason is that there always exist "blind zooms" when using HMD to collect fixations since the participants cannot keep spinning their heads to explore the entire panoptic scene all the time. Consequently, the collected fixations tend to be trapped in some local views, leaving the remaining areas to be the "blind zooms". Therefore, fixation data collected using HMD-based methods that accumulate local views cannot accurately represent the overall global importance of complex panoramic scenes. This paper introduces the auxiliary Window with a Dynamic Blurring (WinDB) fixation collection approach for panoptic video, which doesn't need HMD and is blind-zoom-free. Thus, the collected fixations can well reflect the regional-wise importance degree. Using our WinDB approach, we have released a new PanopticVideo-300 dataset, containing 300 panoptic clips covering over 225 categories. Besides, we have presented a simple baseline design to take full advantage of PanopticVideo-300 to handle the blind-zoom-free attribute-induced fixation shifting problem. Our WinDB approach, PanopticVideo-300, and tailored fixation prediction model are all publicly available at https://github.com/360submit/WinDB.

READ FULL TEXT

page 1

page 2

page 3

page 4

page 5

page 6

page 7

research
03/25/2019

ShopSign: a Diverse Scene Text Dataset of Chinese Shop Signs in Street Views

In this paper, we introduce the ShopSign dataset, which is a newly devel...
research
04/21/2021

NTIRE 2021 Challenge on Quality Enhancement of Compressed Video: Dataset and Study

This paper introduces a novel dataset for video enhancement and studies ...
research
06/29/2023

Foundation Model for Endoscopy Video Analysis via Large-scale Self-supervised Pre-train

Foundation models have exhibited remarkable success in various applicati...
research
05/26/2022

VIDI: A Video Dataset of Incidents

Automatic detection of natural disasters and incidents has become more i...
research
04/01/2019

Passive Head-Mounted Display Music-Listening EEG dataset

We describe the experimental procedures for a dataset that we have made ...
research
10/16/2022

Stochastic Occupancy Grid Map Prediction in Dynamic Scenes

This paper presents two variations of a novel stochastic prediction algo...
research
04/29/2019

TheFragebogen: A Web Browser-based Questionnaire Framework for Scientific Research

Quality of Experience (QoE) typically involves conducting experiments in...

Please sign up or login with your details

Forgot password? Click here to reset