Community detection is a fundamental task in data analysis, and block models provide an approach for identifying a wide variety of community structures while offering high interpretability. The degree-corrected block model (DCBM) is an established model that accounts for the heterogeneity of node degrees. However, inference methods are computationally costly and highly sensitive to initialization, while cheaper alternatives, such as spectral or modularity-based approaches, are restricted to detecting specific structures, typically assortative. In this work, we show that DCBM inference can be reformulated as a constrained nonnegative matrix factorization problem. Leveraging this insight, we propose a novel method for community detection and a theoretically well-grounded initialization strategy that provides an initial estimate of communities for inference algorithms. Our approach is agnostic to any specific network structure and applies to graphs with any structure representable by a DCBM. Experiments on synthetic and real benchmark networks show that our method detects communities comparable to those found by DCBM inference while being faster; for instance, it processes a graph with 100,000 nodes and 1,000,000 edges in approximately 4 minutes. Moreover, the proposed initialization strategy significantly improves solution quality and reduces the number of iterations required by all tested inference algorithms. Overall, this work provides a scalable and robust framework for community detection and highlights the benefits of a matrix-factorization perspective for the DCBM.


翻译:社区检测是数据分析中的基本任务,分块模型为识别多种社区结构提供了方法,同时具有高度可解释性。度修正分块模型(DCBM)是考虑节点度异质性的标准模型。然而,其推断方法计算成本高且对初始化极为敏感,而谱方法或基于模块度的方法等计算成本较低的替代方案局限于检测特定结构(通常是同配结构)。在本文中,我们证明DCBM推断可重新表述为约束非负矩阵分解问题。基于这一见解,我们提出了一种新颖的社区检测方法及一种理论依据充分的初始化策略,可为推断算法提供社区初始估计。该方法不依赖任何特定网络结构,适用于任何可由DCBM表示的图结构。在合成网络和真实基准网络上的实验表明,我们的方法在检测社区方面与DCBM推断结果相当,且速度更快——例如,处理含10万个节点和100万条边的图仅需约4分钟。此外,所提出的初始化策略显著提高了解的质量,并减少了所有测试推断算法所需的迭代次数。总体而言,本文为社区检测提供了一个可扩展且稳健的框架,并凸显了从矩阵分解角度理解DCBM的优势。

0
下载
关闭预览

相关内容

【AAAI2022】基于图神经网络的统一离群点异常检测方法
专知会员服务
28+阅读 · 2022年2月12日
TKDE21 | 网络社团发现新综述:从统计建模到深度学习
专知会员服务
28+阅读 · 2021年10月27日
麦克瑞大学最新「深度学习社区检测」综述论文,28页pdf
异常检测(Anomaly Detection)综述
极市平台
20+阅读 · 2020年10月24日
深度学习时代的目标检测算法
炼数成金订阅号
40+阅读 · 2018年3月19日
深度学习目标检测模型全面综述:Faster R-CNN、R-FCN和SSD
深度学习世界
10+阅读 · 2017年9月18日
侦测欺诈交易(异常点检测)
GBASE数据工程部数据团队
20+阅读 · 2017年5月10日
国家自然科学基金
8+阅读 · 2017年12月31日
国家自然科学基金
0+阅读 · 2017年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
9+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
8+阅读 · 2014年12月31日
VIP会员
最新内容
驱动军事决策变革的顶尖人工智能指挥系统
专知会员服务
1+阅读 · 今天13:50
非对称防御中的自组织临界性:俄乌战争
专知会员服务
8+阅读 · 8月10日
《战争中的大语言模型监管》
专知会员服务
9+阅读 · 8月10日
《边缘计算关键技术分析及美军作战实践应用》
边缘计算的军事应用
专知会员服务
11+阅读 · 8月9日
一种考虑资源机动性的武器目标分配混合算法
专知会员服务
12+阅读 · 8月8日
相关VIP内容
【AAAI2022】基于图神经网络的统一离群点异常检测方法
专知会员服务
28+阅读 · 2022年2月12日
TKDE21 | 网络社团发现新综述:从统计建模到深度学习
专知会员服务
28+阅读 · 2021年10月27日
麦克瑞大学最新「深度学习社区检测」综述论文,28页pdf
相关基金
国家自然科学基金
8+阅读 · 2017年12月31日
国家自然科学基金
0+阅读 · 2017年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
9+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
8+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员