Unsupervised Learning of Sequence Representations by Autoencoders

04/03/2018
by   Wenjie Pei, et al.
0

Traditional machine learning models have problems with handling sequence data, because the lengths of sequences may vary between samples. In this paper, we present an unsupervised learning model for sequence data, called the Integrated Sequence Autoencoder (ISA), to learn a fixed-length vectorial representation by minimizing the reconstruction error. Specifically, we propose to integrate two classical mechanisms for sequence reconstruction which takes into account both the global silhouette information and the local temporal dependencies. Furthermore, we propose a stop feature that serves as a temporal stamp to guide the reconstruction process, and which results in a higher-quality representation. Extensive validation on real-world datasets shows that the learned representation is able to effectively summarize not only the apparent features, but also the underlying and high-level style information. Take for example a speech sequence sample: our ISA model can not only recognize the spoken text (apparent feature), but can also discriminate the speaker who utters the audio (more high-level style).

READ FULL TEXT

page 1

page 2

page 3

page 4

research
03/10/2021

Variable-rate discrete representation learning

Semantically meaningful information content in perceptual signals is usu...
research
06/06/2021

Meta-StyleSpeech : Multi-Speaker Adaptive Text-to-Speech Generation

With rapid progress in neural text-to-speech (TTS) models, personalized ...
research
12/12/2017

auDeep: Unsupervised Learning of Representations from Audio with Deep Recurrent Neural Networks

auDeep is a Python toolkit for deep unsupervised representation learning...
research
10/06/2021

Style Equalization: Unsupervised Learning of Controllable Generative Sequence Models

Controllable generative sequence models with the capability to extract a...
research
10/24/2020

Unsupervised Learning of Disentangled Speech Content and Style Representation

We present an approach for unsupervised learning of speech representatio...
research
07/27/2017

Learning Audio Sequence Representations for Acoustic Event Classification

Acoustic Event Classification (AEC) has become a significant task for ma...
research
01/23/2022

Deep Learning on Attributed Sequences

Recent research in feature learning has been extended to sequence data, ...

Please sign up or login with your details

Forgot password? Click here to reset