Online Self-Attentive Gated RNNs for Real-Time Speaker Separation

06/25/2021
by   Ori Kabeli, et al.
0

Deep neural networks have recently shown great success in the task of blind source separation, both under monaural and binaural settings. Although these methods were shown to produce high-quality separations, they were mainly applied under offline settings, in which the model has access to the full input signal while separating the signal. In this study, we convert a non-causal state-of-the-art separation model into a causal and real-time model and evaluate its performance under both online and offline settings. We compare the performance of the proposed model to several baseline methods under anechoic, noisy, and noisy-reverberant recording conditions while exploring both monaural and binaural inputs and outputs. Our findings shed light on the relative difference between causal and non-causal models when performing separation. Our stateful implementation for online separation leads to a minor drop in performance compared to the offline model; 0.8dB for monaural inputs and 0.3dB for binaural inputs while reaching a real-time factor of 0.65. Samples can be found under the following link: https://kwanum.github.io/sagrnnc-stream-results/.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/03/2022

Conv-NILM-Net, a causal and multi-appliance model for energy source separation

Non-Intrusive Load Monitoring (NILM) seeks to save energy by estimating ...
research
03/14/2023

Towards Real-Time Single-Channel Speech Separation in Noisy and Reverberant Environments

Real-time single-channel speech separation aims to unmix an audio stream...
research
03/03/2021

Compute and memory efficient universal sound source separation

Recent progress in audio source separation lead by deep learning has ena...
research
09/20/2023

Directional Source Separation for Robust Speech Recognition on Smart Glasses

Modern smart glasses leverage advanced audio sensing and machine learnin...
research
10/23/2020

GSEP: A robust vocal and accompaniment separation system using gated CBHG module and loudness normalization

In the field of audio signal processing research, source separation has ...
research
01/10/2022

Noisy Neonatal Chest Sound Separation for High-Quality Heart and Lung Sounds

Stethoscope-recorded chest sounds provide the opportunity for remote car...

Please sign up or login with your details

Forgot password? Click here to reset