Spectrogram Inversion for Audio Source Separation via Consistency, Mixing, and Magnitude Constraints

03/03/2023
by   Paul Magron, et al.
0

Audio source separation is often achieved by estimating the magnitude spectrogram of each source, and then applying a phase recovery (or spectrogram inversion) algorithm to retrieve time-domain signals. Typically, spectrogram inversion is treated as an optimization problem involving one or several terms in order to promote estimates that comply with a consistency property, a mixing constraint, and/or a target magnitude objective. Nonetheless, it is still unclear which set of constraints and problem formulation is the most appropriate in practice. In this paper, we design a general framework for deriving spectrogram inversion algorithm, which is based on formulating optimization problems by combining these objectives either as soft penalties or hard constraints. We solve these by means of algorithms that perform alternating projections on the subsets corresponding to each objective/constraint. Our framework encompasses existing techniques from the literature as well as novel algorithms. We investigate the potential of these approaches for a speech enhancement task. In particular, one of our novel algorithms outperforms other approaches in a realistic setting where the magnitudes are estimated beforehand using a neural network.

READ FULL TEXT
research
10/20/2020

Phase recovery with Bregman divergences for audio source separation

Time-frequency audio source separation is usually achieved by estimating...
research
11/08/2019

Online Spectrogram Inversion for Low-Latency Audio Source Separation

Audio source separation is usually achieved by estimating the short-time...
research
09/30/2016

Phase Unmixing : Multichannel Source Separation with Magnitude Constraints

We consider the problem of estimating the phases of K mixed complex sign...
research
12/28/2011

A general framework for online audio source separation

We consider the problem of online audio source separation. Existing algo...
research
03/23/2021

Learned complex masks for multi-instrument source separation

Music source separation in the time-frequency domain is commonly achieve...
research
11/21/2017

Multichannel Source Separation and Speech Enhancement Using the Convolutive Transfer Function

This paper addresses the problem of audio source recovery from multichan...
research
01/10/2020

Generalized L_p-norm joint inversion of gravity and magnetic data using cross-gradient constraint

A generalized unifying approach for L_p-norm joint inversion of gravity ...

Please sign up or login with your details

Forgot password? Click here to reset