CONVIQT: Contrastive Video Quality Estimator

06/29/2022
by   Pavan C. Madhusudana, et al.
2

Perceptual video quality assessment (VQA) is an integral component of many streaming and video sharing platforms. Here we consider the problem of learning perceptually relevant video quality representations in a self-supervised manner. Distortion type identification and degradation level determination is employed as an auxiliary task to train a deep learning model containing a deep Convolutional Neural Network (CNN) that extracts spatial features, as well as a recurrent unit that captures temporal information. The model is trained using a contrastive loss and we therefore refer to this training framework and resulting model as CONtrastive VIdeo Quality EstimaTor (CONVIQT). During testing, the weights of the trained model are frozen, and a linear regressor maps the learned features to quality scores in a no-reference (NR) setting. We conduct comprehensive evaluations of the proposed model on multiple VQA databases by analyzing the correlations between model predictions and ground-truth quality ratings, and achieve competitive performance when compared to state-of-the-art NR-VQA models, even though it is not trained on those databases. Our ablation experiments demonstrate that the learned representations are highly robust and generalize well across synthetic and realistic distortions. Our results indicate that compelling representations with perceptual bearing can be obtained using self-supervised learning. The implementations used in this work have been made available at https://github.com/pavancm/CONVIQT.

READ FULL TEXT

page 1

page 3

page 4

page 11

research
10/25/2021

Image Quality Assessment using Contrastive Learning

We consider the problem of obtaining image quality representations in a ...
research
07/08/2022

Exploring the Effectiveness of Video Perceptual Representation in Blind Video Quality Assessment

With the rapid growth of in-the-wild videos taken by non-specialists, bl...
research
10/26/2020

ST-GREED: Space-Time Generalized Entropic Differences for Frame Rate Dependent Video Quality Prediction

We consider the problem of conducting frame rate dependent video quality...
research
02/28/2023

Video Quality Assessment with Texture Information Fusion for Streaming Applications

The rise of video streaming applications has increased the demand for Vi...
research
12/12/2018

Power of Tempospatially Unified Spectral Density for Perceptual Video Quality Assessment

We propose a perceptual video quality assessment (PVQA) metric for disto...
research
11/19/2021

DeepQR: Neural-based Quality Ratings for Learnersourced Multiple-Choice Questions

Automated question quality rating (AQQR) aims to evaluate question quali...
research
06/27/2022

Automatic identification of segmentation errors for radiotherapy using geometric learning

Automatic segmentation of organs-at-risk (OARs) in CT scans using convol...

Please sign up or login with your details

Forgot password? Click here to reset