Latent Vector Recovery of Audio GANs

10/16/2020
by   Andrew Keyes, et al.
0

Advanced Generative Adversarial Networks (GANs) are remarkable in generating intelligible audio from a random latent vector. In this paper, we examine the task of recovering the latent vector of both synthesized and real audio. Previous works recovered latent vectors of given audio through an auto-encoder inspired technique that trains an encoder network either in parallel with the GAN or after the generator is trained. With our approach, we train a deep residual neural network architecture to project audio synthesized by WaveGAN into the corresponding latent space with near identical reconstruction performance. To accommodate for the lack of an original latent vector for real audio, we optimize the residual network on the perceptual loss between the real audio samples and the reconstructed audio of the predicted latent vectors. In the case of synthesized audio, the Mean Squared Error (MSE) between the ground truth and recovered latent vector is minimized as well. We further investigated the audio reconstruction performance when several gradient optimization steps are applied to the predicted latent vector. Through our deep neural network based method of training on real and synthesized audio, we are able to predict a latent vector that corresponds to a reasonable reconstruction of real audio. Even though we evaluated our method on WaveGAN, our proposed method is universal and can be applied to any other GANs.

READ FULL TEXT

page 3

page 4

page 5

research
09/11/2020

Inverse mapping of face GANs

Generative adversarial networks (GANs) synthesize realistic images from ...
research
02/15/2017

Precise Recovery of Latent Vectors from Generative Adversarial Networks

Generative adversarial networks (GANs) transform latent vectors into vis...
research
08/23/2021

Adaptable GAN Encoders for Image Reconstruction via Multi-type Latent Vectors with Two-scale Attentions

Although current deep generative adversarial networks (GANs) could synth...
research
03/21/2019

Bandwidth Extension on Raw Audio via Generative Adversarial Networks

Neural network-based methods have recently demonstrated state-of-the-art...
research
10/31/2022

Audio Time-Scale Modification with Temporal Compressing Networks

We proposed a novel approach in the field of time-scale modification on ...
research
09/05/2021

Timbre Transfer with Variational Auto Encoding and Cycle-Consistent Adversarial Networks

This research project investigates the application of deep learning to t...
research
04/06/2021

ReStyle: A Residual-Based StyleGAN Encoder via Iterative Refinement

Recently, the power of unconditional image synthesis has significantly a...

Please sign up or login with your details

Forgot password? Click here to reset