Spectral Clustering with Unbalanced Data

02/20/2013
by   Jing Qian, et al.
0

Spectral clustering (SC) and graph-based semi-supervised learning (SSL) algorithms are sensitive to how graphs are constructed from data. In particular if the data has proximal and unbalanced clusters these algorithms can lead to poor performance on well-known graphs such as k-NN, full-RBF, ϵ-graphs. This is because the objectives such as Ratio-Cut (RCut) or normalized cut (NCut) attempt to tradeoff cut values with cluster sizes, which are not tailored to unbalanced data. We propose a novel graph partitioning framework, which parameterizes a family of graphs by adaptively modulating node degrees in a k-NN graph. We then propose a model selection scheme to choose sizable clusters which are separated by smallest cut values. Our framework is able to adapt to varying levels of unbalancedness of data and can be naturally used for small cluster detection. We theoretically justify our ideas through limit cut analysis. Unsupervised and semi-supervised experiments on synthetic and real data sets demonstrate the superiority of our method.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/07/2012

Graph-based Learning with Unbalanced Clusters

Graph construction is a crucial step in spectral clustering (SC) and gra...
research
05/22/2014

Semi-supervised Spectral Clustering for Classification

We propose a Classification Via Clustering (CVC) algorithm which enables...
research
08/18/2023

A scalable clustering algorithm to approximate graph cuts

Due to its computational complexity, graph cuts for cluster detection an...
research
10/29/2014

Power-Law Graph Cuts

Algorithms based on spectral graph cut objectives such as normalized cut...
research
09/09/2013

Spectral Clustering with Imbalanced Data

Spectral clustering is sensitive to how graphs are constructed from data...
research
12/11/2011

Graph Construction for Learning with Unbalanced Data

Unbalanced data arises in many learning tasks such as clustering of mult...
research
08/26/2016

Clustering and Community Detection with Imbalanced Clusters

Spectral clustering methods which are frequently used in clustering and ...

Please sign up or login with your details

Forgot password? Click here to reset