TMGAN-PLC: Audio Packet Loss Concealment using Temporal Memory Generative Adversarial Network

07/04/2022
by   Yuansheng Guan, et al.
0

Real-time communications in packet-switched networks have become widely used in daily communication, while they inevitably suffer from network delays and data losses in constrained real-time conditions. To solve these problems, audio packet loss concealment (PLC) algorithms have been developed to mitigate voice transmission failures by reconstructing the lost information. Limited by the transmission latency and device memory, it is still intractable for PLC to accomplish high-quality voice reconstruction using a relatively small packet buffer. In this paper, we propose a temporal memory generative adversarial network for audio PLC, dubbed TMGAN-PLC, which is comprised of a novel nested-UNet generator and the time-domain/frequency-domain discriminators. Specifically, a combination of the nested-UNet and temporal feature-wise linear modulation is elaborately devised in the generator to finely adjust the intra-frame information and establish inter-frame temporal dependencies. To complement the missing speech content caused by longer loss bursts, we employ multi-stage gated vector quantizers to capture the correct content and reconstruct the near-real smooth audio. Extensive experiments on the PLC Challenge dataset demonstrate that the proposed method yields promising performance in terms of speech quality, intelligibility, and PLCMOS.

READ FULL TEXT
research
04/11/2022

INTERSPEECH 2022 Audio Deep Packet Loss Concealment Challenge

Audio Packet Loss Concealment (PLC) is the hiding of gaps in audio strea...
research
07/07/2021

Adversarial Auto-Encoding for Packet Loss Concealment

Communication technologies like voice over IP operate under constrained ...
research
11/08/2022

Improving performance of real-time full-band blind packet-loss concealment with predictive network

Packet loss concealment (PLC) is a tool for enhancing speech degradation...
research
05/11/2022

Real-Time Packet Loss Concealment With Mixed Generative and Predictive Model

As deep speech enhancement algorithms have recently demonstrated capabil...
research
04/04/2022

tPLCnet: Real-time Deep Packet Loss Concealment in the Time Domain Using a Short Temporal Context

This paper introduces a real-time time-domain packet loss concealment (P...
research
07/03/2022

Towards Error-Resilient Neural Speech Coding

Neural audio coding has shown very promising results recently in the lit...
research
05/22/2019

Effects of Packet Loss and Jitter on VoLTE Call Quality

This work performs a preliminary, comparative analysis of the end-to-end...

Please sign up or login with your details

Forgot password? Click here to reset