A Convolutional Neural Network model based on Neutrosophy for Noisy Speech Recognition

01/27/2019
by   Elyas Rashno, et al.
0

Convolutional neural networks are sensitive to unknown noisy condition in the test phase and so their performance degrades for the noisy data classification task including noisy speech recognition. In this research, a new convolutional neural network (CNN) model with data uncertainty handling; referred as NCNN (Neutrosophic Convolutional Neural Network); is proposed for classification task. Here, speech signals are used as input data and their noise is modeled as uncertainty. In this task, using speech spectrogram, a definition of uncertainty is proposed in neutrosophic (NS) domain. Uncertainty is computed for each Time-frequency point of speech spectrogram as like a pixel. Therefore, uncertainty matrix with the same size of spectrogram is created in NS domain. In the next step, a two parallel paths CNN classification model is proposed. Speech spectrogram is used as input of the first path and uncertainty matrix for the second path. The outputs of two paths are combined to compute the final output of the classifier. To show the effectiveness of the proposed method, it has been compared with conventional CNN on the isolated words of Aurora2 dataset. The proposed method achieves the average accuracy of 85.96 in noisy train data. It is more robust against Car, Airport and Subway noises with accuracies 90, 88 and 81 in test sets A, B and C, respectively. Results show that the proposed method outperforms conventional CNN with the improvement of 6, 5 and 2 percentage in test set A, test set B and test sets C, respectively. It means that the proposed method is more robust against noisy data and handle these data effectively.

READ FULL TEXT

page 1

page 3

research
01/22/2022

A Noise-Robust Self-supervised Pre-training Model Based Speech Representation Learning for Automatic Speech Recognition

Wav2vec2.0 is a popular self-supervised pre-training framework for learn...
research
08/09/2021

Time-Frequency Localization Using Deep Convolutional Maxout Neural Network in Persian Speech Recognition

In this paper, a CNN-based structure for the time-frequency localization...
research
02/03/2021

Downbeat Tracking with Tempo-Invariant Convolutional Neural Networks

The human ability to track musical downbeats is robust to changes in tem...
research
04/12/2019

A robust approach to model-based classification based on trimming and constraints

In a standard classification framework a set of trustworthy learning dat...
research
05/16/2023

Noise robust neural network architecture

In which we propose neural network architecture (dune neural network) fo...
research
06/03/2022

Distributional loss for convolutional neural network regression and application to GNSS multi-path estimation

Convolutional Neural Network (CNN) have been widely used in image classi...
research
07/03/2020

Balanced Symmetric Cross Entropy for Large Scale Imbalanced and Noisy Data

Deep convolution neural network has attracted many attentions in large-s...

Please sign up or login with your details

Forgot password? Click here to reset