Multi-stream Network With Temporal Attention For Environmental Sound Classification

01/24/2019
by   Xinyu Li, et al.
0

Environmental sound classification systems often do not perform robustly across different sound classification tasks and audio signals of varying temporal structures. We introduce a multi-stream convolutional neural network with temporal attention that addresses these problems. The network relies on three input streams consisting of raw audio and spectral features and utilizes a temporal attention function computed from energy changes over time. Training and classification utilizes decision fusion and data augmentation techniques that incorporate uncertainty. We evaluate this network on three commonly used data sets for environmental sound and audio scene classification and achieve new state-of-the-art performance without any changes in network architecture or front-end preprocessing, thus demonstrating better generalizability.

READ FULL TEXT
research
08/15/2016

Deep Convolutional Neural Networks and Data Augmentation for Environmental Sound Classification

The ability of deep convolutional neural networks (CNN) to learn discrim...
research
09/15/2023

SSL-Net: A Synergistic Spectral and Learning-based Network for Efficient Bird Sound Classification

Efficient and accurate bird sound classification is of important for eco...
research
05/24/2018

Environmental Sound Classification Based on Multi-temporal Resolution Convolutional Neural Network Combining with Multi-level Features

Motivated by the fact that characteristics of different sound classes ar...
research
05/24/2018

Environmental Sound Classification Based on Multi-temporal Resolution CNN Network Combining with Multi-level Features

Motivated by the fact that characteristics of different sound classes ar...
research
08/20/2019

AI for Earth: Rainforest Conservation by Acoustic Surveillance

Saving rainforests is a key to halting adverse climate changes. In this ...
research
04/06/2019

Spatio-Temporal Attention Pooling for Audio Scene Classification

Acoustic scenes are rich and redundant in their content. In this work, w...
research
11/21/2019

An End-to-End Audio Classification System based on Raw Waveforms and Mix-Training Strategy

Audio classification can distinguish different kinds of sounds, which is...

Please sign up or login with your details

Forgot password? Click here to reset