A Robust Frame-based Nonlinear Prediction System for Automatic Speech Coding

01/22/2016
by   Mahmood Yousefi-Azar, et al.
0

In this paper, we propose a neural-based coding scheme in which an artificial neural network is exploited to automatically compress and decompress speech signals by a trainable approach. Having a two-stage training phase, the system can be fully specified to each speech frame and have robust performance across different speakers and wide range of spoken utterances. Indeed, Frame-based nonlinear predictive coding (FNPC) would code a frame in the procedure of training to predict the frame samples. The motivating objective is to analyze the system behavior in regenerating not only the envelope of spectra, but also the spectra phase. This scheme has been evaluated in time and discrete cosine transform (DCT) domains and the output of predicted phonemes show the potentiality of the FNPC to reconstruct complicated signals. The experiments were conducted on three voiced plosive phonemes, b/d/g/ in time and DCT domains versus the number of neurons in the hidden layer. Experiments approve the FNPC capability as an automatic coding system by which /b/d/g/ phonemes have been reproduced with a good accuracy. Evaluations revealed that the performance of FNPC system, trained to predict DCT coefficients is more desirable, particularly for frames with the wider distribution of energy, compared to time samples.

READ FULL TEXT

page 8

page 9

page 10

research
04/04/2022

Nonlinear Vectorial Prediction with Neural Nets

In this paper we propose a nonlinear vectorial prediction scheme based o...
research
03/03/2022

Nonlinear predictive models computation in ADPCM schemes

Recently several papers have been published on nonlinear prediction appl...
research
08/17/2023

Long-frame-shift Neural Speech Phase Prediction with Spectral Continuity Enhancement and Interpolation Error Compensation

Speech phase prediction, which is a significant research focus in the fi...
research
04/11/2020

Improved Speech Representations with Multi-Target Autoregressive Predictive Coding

Training objectives based on predictive coding have recently been shown ...
research
03/31/2022

A comparative study between linear and nonlinear speech prediction

This paper is focused on nonlinear prediction coding, which consists on ...
research
04/01/2022

Adaptive hybrid speech coding with a MLP LPC structure

In the last years there has been a growing interest for nonlinear speech...

Please sign up or login with your details

Forgot password? Click here to reset