DCASE 2018 Challenge Surrey Cross-Task convolutional neural network baseline

08/02/2018
by   Qiuqiang Kong, et al.
0

The Detection and Classification of Acoustic Scenes and Events (DCASE) consists of five audio classification and sound event detection tasks: 1) Acoustic scene classification, 2) General-purpose audio tagging of Freesound, 3) Bird audio detection, 4) Weakly-labeled semi-supervised sound event detection and 5) Multi-channel audio classification. In this paper, we create a cross-task baseline system for all five tasks based on a convlutional neural network (CNN): a `CNN Baseline' system. We implemented CNNs with 4 layers and 8 layers originating from AlexNet and VGG from computer vision. We investigated how the performance varies from task to task with the same configuration of neural networks. Experiments show that deeper CNN with 8 layers performs better than CNN with 4 layers on all tasks except Task 1. Using CNN with 8 layers, we achieve an accuracy of 0.680 on Task 1, an accuracy of 0.895 and a mean average precision (MAP) of 0.928 on Task 2, an accuracy of 0.751 and an area under the curve (AUC) of 0.854 on Task 3, a sound event detection F1 score of 20.8% on Task 4, and an F1 score of 87.75% on Task 5. We released the Python source code of the baseline systems under the MIT liscense for further research.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/02/2018

DCASE 2018 Challenge baseline with convolutional neural networks

The Detection and Classification of Acoustic Scenes and Events (DCASE) i...
research
02/28/2023

Incremental Learning of Acoustic Scenes and Sound Events

In this paper, we propose a method for incremental learning of two disti...
research
07/28/2021

Deep learning based cough detection camera using enhanced features

Coughing is a typical symptom of COVID-19. To detect and localize coughi...
research
08/17/2021

Neonatal Bowel Sound Detection Using Convolutional Neural Network and Laplace Hidden Semi-Markov Model

Abdominal auscultation is a convenient, safe and inexpensive method to a...
research
04/06/2019

Cross-task learning for audio tagging, sound event detection and spatial localization: DCASE 2019 baseline systems

The Detection and Classification of Acoustic Scenes and Events (DCASE) 2...
research
01/22/2020

Non-Negative Matrix Factorization-Convolutional Neural Network (NMF-CNN) For Sound Event Detection

The main scientific question of this year DCASE challenge, Task 4 - Soun...
research
07/30/2018

DCASE 2018 Challenge - Task 5: Monitoring of domestic activities based on multi-channel acoustics

The DCASE 2018 Challenge consists of five tasks related to automatic cla...

Please sign up or login with your details

Forgot password? Click here to reset