Deep Learning-Based Single-Ended Objective Quality Measures for Time-Scale Modified Audio

09/07/2020
by   Timothy Roberts, et al.
0

Objective evaluation of audio processed with Time-Scale Modification (TSM) is seeing a resurgence of interest. Recently, a labelled time-scaled audio dataset was used to train an objective measure for TSM evaluation. This DE measure was an extension of Perceptual Evaluation of Audio Quality, and required reference and test signals. In this paper, two single-ended objective quality measures for time-scaled audio are proposed that do not require a reference signal. Data driven features are created by either a convolutional neural network (CNN) or a bidirectional gated recurrent unit (BGRU) network and fed to a fully-connected network to predict subjective mean opinion scores. The proposed CNN and BGRU measures achieve an average Root Mean Squared Error of 0.608 and 0.576, and a mean Pearson correlation of 0.771 and 0.794, respectively. The proposed measures are used to evaluate TSM algorithms, and comparisons are provided for 16 TSM implementations. The objective measure is available at https://www.github.com/zygurt/TSM.

READ FULL TEXT
research
06/11/2020

An Objective Measure of Quality for Time-Scale Modification of Audio

Objective evaluation of audio processed with Time-Scale Modification (TS...
research
03/16/2019

Non-intrusive speech quality assessment using neural networks

Estimating the perceived quality of an audio signal is critical for many...
research
06/01/2020

A time-scale modification dataset with subjective quality labels

Time Scale Modification (TSM) is a well-researched field; however, no ef...
research
10/31/2022

Audio Time-Scale Modification with Temporal Compressing Networks

We proposed a novel approach in the field of time-scale modification on ...
research
10/21/2021

Objective Measures of Perceptual Audio Quality Reviewed: An Evaluation of Their Application Domain Dependence

Over the past few decades, computational methods have been developed to ...
research
06/13/2023

Evaluation of Spatial Distortion in Multichannel Audio

Despite the recent proliferation of spatial audio technologies, the eval...
research
09/19/2019

WEnets: A Convolutional Framework for Evaluating Audio Waveforms

We describe a new convolutional framework for waveform evaluation, WEnet...

Please sign up or login with your details

Forgot password? Click here to reset