Attentive Representation Learning with Adversarial Training for Short Text Clustering

12/08/2019
by   Wei Zhang, et al.
0

Short text clustering has far-reaching effects on semantic analysis, showing its importance for multiple applications such as corpus summarization and information retrieval. However, it inevitably encounters the severe sparsity of short text representation, making the previous clustering approaches still far from satisfactory. In this paper, we present a novel attentive representation learning model for shot text clustering, wherein cluster-level attention is proposed to capture the correlation between text representation and cluster representation. Relying on this, the representation learning and clustering for short text are seamlessly integrated into a unified framework. To further facilitate the model training process, we apply adversarial training to the unsupervised clustering setting, by adding perturbations to the cluster representations. The model parameters and perturbations are optimized alternately through a minimax game. Extensive experiments on three real-world short text datasets demonstrate the superiority of the proposed model over several strong competitors, verifying that adversarial training yields a substantial performance gain.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
09/14/2021

Improving Gradient-based Adversarial Training for Text Classification by Contrastive Learning and Auto-Encoder

Recent work has proposed several efficient approaches for generating gra...
research
06/18/2019

Mimicking Human Process: Text Representation via Latent Semantic Clustering for Classification

Considering that words with different characteristic in the text have di...
research
09/21/2021

Representation Learning for Short Text Clustering

Effective representation learning is critical for short text clustering ...
research
03/22/2019

An end-to-end Neural Network Framework for Text Clustering

The unsupervised text clustering is one of the major tasks in natural la...
research
10/03/2016

Nonsymbolic Text Representation

We introduce the first generic text representation model that is complet...
research
10/29/2020

Attentive Clustering Processes

Amortized approaches to clustering have recently received renewed attent...
research
06/18/2020

Online Deep Clustering for Unsupervised Representation Learning

Joint clustering and feature learning methods have shown remarkable perf...

Please sign up or login with your details

Forgot password? Click here to reset