Unsupervised Embedding of Hierarchical Structure in Euclidean Space

10/30/2020
by   Jinyu Zhao, et al.
0

Deep embedding methods have influenced many areas of unsupervised learning. However, the best methods for learning hierarchical structure use non-Euclidean representations, whereas Euclidean geometry underlies the theory behind many hierarchical clustering algorithms. To bridge the gap between these two areas, we consider learning a non-linear embedding of data into Euclidean space as a way to improve the hierarchical clustering produced by agglomerative algorithms. To learn the embedding, we revisit using a variational autoencoder with a Gaussian mixture prior, and we show that rescaling the latent space embedding and then applying Ward's linkage-based algorithm leads to improved results for both dendrogram purity and the Moseley-Wang cost function. Finally, we complement our empirical results with a theoretical explanation of the success of this approach. We study a synthetic model of the embedded vectors and prove that Ward's method exactly recovers the planted hierarchical clustering with high probability.

READ FULL TEXT

page 8

page 24

research
10/29/2019

Hyperbolic Node Embedding for Signed Networks

The rapid evolving World Wide Web has produced a large amount of complex...
research
09/24/2021

Non-Euclidean Self-Organizing Maps

Self-Organizing Maps (SOMs, Kohonen networks) belong to neural network m...
research
05/28/2019

Variational Information Bottleneck for Unsupervised Clustering: Deep Gaussian Mixture Embedding

In this paper, we develop an unsupervised generative clustering framewor...
research
12/25/2020

Learning Robust Representation for Clustering through Locality Preserving Variational Discriminative Network

Clustering is one of the fundamental problems in unsupervised learning. ...
research
08/07/2023

Wide Gaps and Clustering Axioms

The widely applied k-means algorithm produces clusterings that violate o...
research
12/07/2020

Joint Optimization of an Autoencoder for Clustering and Embedding

Incorporating k-means-like clustering techniques into (deep) autoencoder...
research
03/14/2022

Unsupervised Clustering of Roman Potsherds via Variational Autoencoders

In this paper we propose an artificial intelligence imaging solution to ...

Please sign up or login with your details

Forgot password? Click here to reset