Community Detection in the Hypergraph SBM: Optimal Recovery Given the Similarity Matrix

08/23/2022
by   Julia Gaudio, et al.
0

Community detection is a fundamental problem in network science. In this paper, we consider community detection in hypergraphs drawn from the hypergraph stochastic block model (HSBM), with a focus on exact community recovery. We study the performance of polynomial-time algorithms for community detection in a case where the full hypergraph is unknown. Instead, we are provided a similarity matrix W, where W_ij reports the number of hyperedges containing both i and j. Under this information model, Kim, Bandeira, and Goemans [KBG18] determined the information-theoretic threshold for exact recovery, and proposed a semidefinite programming relaxation which they conjectured to be optimal. In this paper, we confirm this conjecture. We also show that a simple, highly efficient spectral algorithm is optimal, establishing the spectral algorithm as the method of choice. Our analysis of the spectral algorithm crucially relies on strong entrywise bounds on the eigenvectors of W. Our bounds are inspired by the work of Abbe, Fan, Wang, and Zhong [AFWZ20], who developed entrywise bounds for eigenvectors of symmetric matrices with independent entries. Despite the complex dependency structure in similarity matrices, we prove similar entrywise guarantees.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
11/04/2021

Community detection in censored hypergraph

Community detection refers to the problem of clustering the nodes of a n...
research
04/11/2019

Community Detection in the Sparse Hypergraph Stochastic Block Model

We consider the community detection problem in sparse random hypergraphs...
research
03/14/2022

Sparse random hypergraphs: Non-backtracking spectra and community detection

We consider the community detection problem in a sparse q-uniform hyperg...
research
11/23/2020

Statistical and computational thresholds for the planted k-densest sub-hypergraph problem

Recovery a planted signal perturbed by noise is a fundamental problem in...
research
12/28/2021

Non-Convex Joint Community Detection and Group Synchronization via Generalized Power Method

This paper proposes a Generalized Power Method (GPM) to tackle the probl...
research
05/24/2017

Provable Estimation of the Number of Blocks in Block Models

Community detection is a fundamental unsupervised learning problem for u...
research
01/27/2023

Multilayer hypergraph clustering using the aggregate similarity matrix

We consider the community recovery problem on a multilayer variant of th...

Please sign up or login with your details

Forgot password? Click here to reset