Audio Time-Scale Modification with Temporal Compressing Networks

10/31/2022
by   Ernie Chu, et al.
0

We proposed a novel approach in the field of time-scale modification on audio signals. While traditional methods use the framing technique, spectral approach uses the short-time Fourier transform to preserve the frequency during temporal stretching. TSM-Net, our neural-network model encodes the raw audio into a high-level latent representation. We call it Neuralgram, in which one vector represents 1024 audio samples. It is inspired by the framing technique but addresses the clipping artifacts. The Neuralgram is a two-dimensional matrix with real values, we can apply some existing image resizing techniques on the Neuralgram and decode it using our neural decoder to obtain the time-scaled audio. Both the encoder and decoder are trained with GANs, which shows fair generalization ability on the scaled Neuralgrams. Our method yields little artifacts and opens a new possibility in the research of modern time-scale modification. The audio samples can be found on https://ernestchu.github.io/tsm-net-demo/

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/11/2020

An Objective Measure of Quality for Time-Scale Modification of Audio

Objective evaluation of audio processed with Time-Scale Modification (TS...
research
09/07/2020

Deep Learning-Based Single-Ended Objective Quality Measures for Time-Scale Modified Audio

Objective evaluation of audio processed with Time-Scale Modification (TS...
research
07/21/2021

Audio Captioning Transformer

Audio captioning aims to automatically generate a natural language descr...
research
10/16/2020

Latent Vector Recovery of Audio GANs

Advanced Generative Adversarial Networks (GANs) are remarkable in genera...
research
07/13/2022

Masked Autoencoders that Listen

This paper studies a simple extension of image-based Masked Autoencoders...
research
10/24/2022

High Fidelity Neural Audio Compression

We introduce a state-of-the-art real-time, high-fidelity, audio codec le...
research
04/25/2023

AI-Synthesized Voice Detection Using Neural Vocoder Artifacts

Advancements in AI-synthesized human voices have created a growing threa...

Please sign up or login with your details

Forgot password? Click here to reset