DOANet: a deep dilated convolutional neural network approach for search and rescue with drone-embedded sound source localization

03/02/2021
by   K M Naimul Hassan, et al.
0

Drone-embedded sound source localization (SSL) has interesting application perspective in challenging search and rescue scenarios due to bad lighting conditions or occlusions. However, the problem gets complicated by severe drone ego-noise that may result in negative signal-to-noise ratios in the recorded microphone signals. In this paper, we present our work on drone-embedded SSL using recordings from an 8-channel cube-shaped microphone array embedded in an unmanned aerial vehicle (UAV). We use angular spectrum-based TDOA (time difference of arrival) estimation methods such as generalized cross-correlation phase-transform (GCC-PHAT), minimum-variance-distortion-less-response (MVDR) as baseline, which are state-of-the-art techniques for SSL. Though we improve the baseline method by reducing ego-noise using speed correlated harmonics cancellation (SCHC) technique, our main focus is to utilize deep learning techniques to solve this challenging problem. Here, we propose an end-to-end deep learning model, called DOANet, for SSL. DOANet is based on a one-dimensional dilated convolutional neural network that computes the azimuth and elevation angles of the target sound source from the raw audio signal. The advantage of using DOANet is that it does not require any hand-crafted audio features or ego-noise reduction for DOA estimation. We then evaluate the SSL performance using the proposed and baseline methods and find that the DOANet shows promising results compared to both the angular spectrum methods with and without SCHC. To evaluate the different methods, we also introduce a well-known parameter—area under the curve (AUC) of cumulative histogram plots of angular deviations—as a performance indicator which, to our knowledge, has not been used as a performance indicator for this sort of problem before.

READ FULL TEXT

page 1

page 6

page 8

research
01/17/2021

An embedded multichannel sound acquisition system for drone audition

Microphone array techniques can improve the acoustic sensing performance...
research
07/29/2018

Towards End-to-End Acoustic Localization using Deep Learning: from Audio Signal to Source Position Coordinates

This paper presents a novel approach for indoor acoustic source localiza...
research
07/03/2019

Audio-Based Search and Rescue with a Drone: Highlights from the IEEE Signal Processing Cup 2019 Student Competition

Unmanned aerial vehicles (UAV), commonly referred to as drones, have rai...
research
11/04/2022

Speech enhancement using ego-noise references with a microphone array embedded in an unmanned aerial vehicle

A method is proposed for performing speech enhancement using ego-noise r...
research
09/08/2021

A Survey of Sound Source Localization with Deep Learning Methods

This article is a survey on deep learning methods for single and multipl...
research
08/05/2021

SLoClas: A Database for Joint Sound Localization and Classification

In this work, we present the development of a new database, namely Sound...
research
04/23/2023

Sound-based drone fault classification using multitask learning

The drone has been used for various purposes, including military applica...

Please sign up or login with your details

Forgot password? Click here to reset