An Objective for Hierarchical Clustering in Euclidean Space and its Connection to Bisecting K-means

08/30/2020
by   Benjamin Moseley, et al.
13

This paper explores hierarchical clustering in the case where pairs of points have dissimilarity scores (e.g. distances) as a part of the input. The recently introduced objective for points with dissimilarity scores results in every tree being a 1/2 approximation if the distances form a metric. This shows the objective does not make a significant distinction between a good and poor hierarchical clustering in metric spaces. Motivated by this, the paper develops a new global objective for hierarchical clustering in Euclidean space. The objective captures the criterion that has motivated the use of divisive clustering algorithms: that when a split happens, points in the same cluster should be more similar than points in different clusters. Moreover, this objective gives reasonable results on ground-truth inputs for hierarchical clustering. The paper builds a theoretical connection between this objective and the bisecting k-means algorithm. This paper proves that the optimal 2-means solution results in a constant approximation for the objective. This is the first paper to show the bisecting k-means algorithm optimizes a natural global objective over the entire tree.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/02/2023

Multi Layer Peeling for Linear Arrangement and Hierarchical Clustering

We present a new multi-layer peeling technique to cluster points in a me...
research
12/31/2019

Scalable Hierarchical Clustering with Tree Grafting

We introduce Grinch, a new algorithm for large-scale, non-greedy hierarc...
research
09/03/2021

Stability for layer points

In the first half this paper, we generalize the theory of layer points f...
research
06/26/2023

Minimum Description Length Clustering to Measure Meaningful Image Complexity

Existing image complexity metrics cannot distinguish meaningful content ...
research
07/31/2017

Temporal Hierarchical Clustering

We study hierarchical clusterings of metric spaces that change over time...
research
11/08/2019

Convex Hierarchical Clustering for Graph-Structured Data

Convex clustering is a recent stable alternative to hierarchical cluster...

Please sign up or login with your details

Forgot password? Click here to reset