Unsupervised Feature Learning for Environmental Sound Classification Using Cycle Consistent Generative Adversarial Network

04/08/2019
by   Mohammad Esmaeilpour, et al.
0

In this paper we propose a novel environmental sound classification approach incorporating unsupervised feature learning from codebook via spherical K-Means++ algorithm and a new architecture for high-level data augmentation. The audio signal is transformed into a 2D representation using a discrete wavelet transform (DWT). The DWT spectrograms are then augmented by a novel architecture for cycle-consistent generative adversarial network. This high-level augmentation bootstraps generated spectrograms in both intra and inter class manners by translating structural features from sample to sample. A codebook is built by coding the DWT spectrograms with the speeded-up robust feature detector (SURF) and the K-Means++ algorithm. The Random Forest is our final learning algorithm which learns the environmental sound classification task from the clustered codewords in the codebook. Experimental results in four benchmarking environmental sound datasets (ESC-10, ESC-50, UrbanSound8k, and DCASE-2017) have shown that the proposed classification approach outperforms the state-of-the-art classifiers in the scope, including advanced and dense convolutional neural networks such as AlexNet and GoogLeNet, improving the classification rate between 3.51

READ FULL TEXT
research
08/15/2016

Deep Convolutional Neural Networks and Data Augmentation for Environmental Sound Classification

The ability of deep convolutional neural networks (CNN) to learn discrim...
research
04/18/2019

End-to-End Environmental Sound Classification using a 1D Convolutional Neural Network

In this paper, we present an end-to-end approach for environmental sound...
research
04/15/2021

EnvGAN: Adversarial Synthesis of Environmental Sounds for Data Augmentation

The research in Environmental Sound Classification (ESC) has been progre...
research
03/27/2023

Data Augmentation for Environmental Sound Classification Using Diffusion Probabilistic Model with Top-k Selection Discriminator

Despite consistent advancement in powerful deep learning techniques in r...
research
07/27/2020

From Sound Representation to Model Robustness

In this paper, we demonstrate the extreme vulnerability of a residual de...
research
06/29/2022

DrumGAN VST: A Plugin for Drum Sound Analysis/Synthesis With Autoencoding Generative Adversarial Networks

In contemporary popular music production, drum sound design is commonly ...
research
05/25/2018

Masked Conditional Neural Networks for Environmental Sound Classification

The ConditionaL Neural Network (CLNN) exploits the nature of the tempora...

Please sign up or login with your details

Forgot password? Click here to reset