Semi-Supervised Learning with Data Augmentation for End-to-End ASR

07/27/2020
by   Felix Weninger, et al.
0

In this paper, we apply Semi-Supervised Learning (SSL) along with Data Augmentation (DA) for improving the accuracy of End-to-End ASR. We focus on the consistency regularization principle, which has been successfully applied to image classification tasks, and present sequence-to-sequence (seq2seq) versions of the FixMatch and Noisy Student algorithms. Specifically, we generate the pseudo labels for the unlabeled data on-the-fly with a seq2seq model after perturbing the input features with DA. We also propose soft label variants of both algorithms to cope with pseudo label errors, showing further performance improvements. We conduct SSL experiments on a conversational speech data set with 1.9kh manually transcribed training data, using only 25 labels (475h labeled data). In the result, the Noisy Student algorithm with soft labels and consistency regularization achieves 10.4 reduction when adding 475h of unlabeled data, corresponding to a recovery rate of 92 SSL performance is within 5 training set (recovery rate: 78

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/14/2022

Improved Consistency Training for Semi-Supervised Sequence-to-Sequence ASR via Speech Chain Reconstruction and Self-Transcribing

Consistency regularization has recently been applied to semi-supervised ...
research
11/30/2022

Semi-Supervised Heterogeneous Graph Learning with Multi-level Data Augmentation

In recent years, semi-supervised graph learning with data augmentation (...
research
01/24/2020

Semi-supervised ASR by End-to-end Self-training

While deep learning based end-to-end automatic speech recognition (ASR) ...
research
11/11/2022

Continuous Soft Pseudo-Labeling in ASR

Continuous pseudo-labeling (PL) algorithms such as slimIPL have recently...
research
03/28/2020

Gradient-based Data Augmentation for Semi-Supervised Learning

In semi-supervised learning (SSL), a technique called consistency regula...
research
09/13/2023

Reliability-based cleaning of noisy training labels with inductive conformal prediction in multi-modal biomedical data mining

Accurately labeling biomedical data presents a challenge. Traditional se...
research
06/12/2019

Manifold Graph with Learned Prototypes for Semi-Supervised Image Classification

Recent advances in semi-supervised learning methods rely on estimating c...

Please sign up or login with your details

Forgot password? Click here to reset