Subsampling-Based Modified Bayesian Information Criterion for Large-Scale Stochastic Block Models

04/14/2023
by   Jiayi Deng, et al.
0

Identifying the number of communities is a fundamental problem in community detection, which has received increasing attention recently. However, rapid advances in technology have led to the emergence of large-scale networks in various disciplines, thereby making existing methods computationally infeasible. To address this challenge, we propose a novel subsampling-based modified Bayesian information criterion (SM-BIC) for identifying the number of communities in a network generated via the stochastic block model and degree-corrected stochastic block model. We first propose a node-pair subsampling method to extract an informative subnetwork from the entire network, and then we derive a purely data-driven criterion to identify the number of communities for the subnetwork. In this way, the SM-BIC can identify the number of communities based on the subsampled network instead of the entire dataset. This leads to important computational advantages over existing methods. We theoretically investigate the computational complexity and identification consistency of the SM-BIC. Furthermore, the advantages of the SM-BIC are demonstrated by extensive numerical studies.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/12/2022

Multiple Hypothesis Testing To Estimate The Number of Communities in Sparse Stochastic Block Models

Network-based clustering methods frequently require the number of commun...
research
11/11/2020

A Distributed Algorithm for Overlapped Community Detection in Large-Scale Networks

Overlapped community detection in social networks has become an importan...
research
01/16/2014

Community Detection in Networks using Graph Distance

The study of networks has received increased attention recently not only...
research
06/06/2018

Identifying Heritable Communities of Microbiome by Root-Unifrac and Wishart Distribution

We introduce a method to identify heritable microbiome communities when ...
research
09/04/2018

Determining the Number of Communities in Degree-corrected Stochastic Block Models

We propose to estimate the number of communities in degree-corrected sto...
research
12/04/2014

How Many Communities Are There?

Stochastic blockmodels and variants thereof are among the most widely us...
research
12/31/2022

Efficient Methods for Approximating the Shapley Value for Asset Sharing in Energy Communities

With the emergence of energy communities, where a number of prosumers in...

Please sign up or login with your details

Forgot password? Click here to reset