Graph pattern counting serves as a cornerstone of network analysis with extensive real-world applications. Its integration with local differential privacy (LDP) has gained growing attention for protecting sensitive graph information in decentralized settings. However, existing LDP frameworks are largely ad hoc, offering solutions only for specific patterns such as triangles and stars. A general mechanism for counting arbitrary graph patterns, even for the subclass of acyclic patterns, has remained an open problem. To fill this gap, we present the first general solution for counting arbitrary acyclic patterns under LDP. We identify and tackle two fundamental challenges: generalizing pattern construction from distributed data and eliminating node duplication during the construction. To address the first challenge, we propose an LDP-tailored recursive subpattern counting framework that incrementally builds patterns across multiple communication rounds. For the second challenge, we apply a random marking technique that restricts each node to a unique position in the pattern during computation. Our mechanism achieves strong utility guarantees: for any acyclic graph pattern with $k$ edges, we achieve an additive error of $\tilde{O}(\sqrt{N}d(G)^k)$, where $N$ is the number of nodes and $d(G)$ is the maximum degree of the input graph $G$. Experiments on real-world graph datasets across multiple types of acyclic patterns demonstrate that our mechanisms achieve up to $46$-$2600\times$ improvement in utility and $300$-$650\times$ reduction in communication cost compared to the baseline methods.


翻译:图模式计数作为网络分析的基石,在现实世界中有广泛的应用。它与局部差分隐私(LDP)的结合因能在去中心化场景中保护敏感图信息而日益受到关注。然而,现有的LDP框架大多具有特殊性,仅针对特定模式(如三角形和星形)提供解决方案。即使对于无环模式子类,对任意图模式进行计数的通用机制仍然是一个未解决的问题。为填补这一空白,我们提出了首个在LDP下对任意无环模式进行计数的通用解决方案。我们识别并解决了两个基本挑战:从分布式数据中泛化模式构建,以及在此构建过程中消除节点重复。针对第一个挑战,我们提出了一种面向LDP的递归子模式计数框架,该框架通过多轮通信逐步构建模式。针对第二个挑战,我们应用了一种随机标记技术,在计算过程中将每个节点限制到模式中的唯一位置。我们的机制实现了强大的效用保证:对于任意包含$k$条边的无环图模式,我们实现了$\tilde{O}(\sqrt{N}d(G)^k)$的加性误差,其中$N$是节点数,$d(G)$是输入图$G$的最大度数。在多种无环模式类型的真实图数据集上的实验表明,与基线方法相比,我们的机制在效用上实现了高达$46$-$2600$倍的提升,在通信开销上减少了$300$-$650$倍。

0
下载
关闭预览

相关内容

【斯坦福博士论文】有效的差分隐私深度学习,153页pdf
专知会员服务
19+阅读 · 2024年7月10日
【2024新书】数据科学中的图算法:以Neo4j为例
专知会员服务
81+阅读 · 2024年1月19日
图数据上的隐私攻击与防御技术
专知会员服务
28+阅读 · 2022年4月28日
专知会员服务
49+阅读 · 2021年8月1日
专知会员服务
52+阅读 · 2021年6月16日
专知会员服务
45+阅读 · 2020年12月26日
专知会员服务
52+阅读 · 2020年12月10日
专知会员服务
41+阅读 · 2020年12月1日
【图计算】人工智能之图计算
产业智能官
17+阅读 · 2020年4月3日
图机器学习 2.2-2.4 Properties of Networks, Random Graph
图与推荐
10+阅读 · 2020年3月28日
CVPR 2019 | 无监督领域特定单图像去模糊
PaperWeekly
14+阅读 · 2019年3月20日
图分类:结合胶囊网络Capsule和图卷积GCN(附代码)
中国人工智能学会
36+阅读 · 2019年2月26日
图神经网络最近这么火,不妨看看我们精选的这七篇
人工智能前沿讲习班
37+阅读 · 2018年12月10日
超像素、语义分割、实例分割、全景分割 傻傻分不清?
计算机视觉life
19+阅读 · 2018年11月27日
国家自然科学基金
0+阅读 · 2017年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
Arxiv
0+阅读 · 4月2日
VIP会员
相关主题
最新内容
俄乌无人机战争的六大启示
专知会员服务
9+阅读 · 8月3日
《无人机空中监控:通信实验洞察》
专知会员服务
6+阅读 · 8月3日
从采集到决策:美军视角下的战术情报范式重构
《履带式无人地面战车技术发展现状》
专知会员服务
6+阅读 · 8月2日
《无人机脆弱性利用:网络空间力量的新域》
专知会员服务
9+阅读 · 8月1日
美空军如何将人工智能从战场部署至后方机关
专知会员服务
14+阅读 · 7月31日
相关VIP内容
【斯坦福博士论文】有效的差分隐私深度学习,153页pdf
专知会员服务
19+阅读 · 2024年7月10日
【2024新书】数据科学中的图算法:以Neo4j为例
专知会员服务
81+阅读 · 2024年1月19日
图数据上的隐私攻击与防御技术
专知会员服务
28+阅读 · 2022年4月28日
专知会员服务
49+阅读 · 2021年8月1日
专知会员服务
52+阅读 · 2021年6月16日
专知会员服务
45+阅读 · 2020年12月26日
专知会员服务
52+阅读 · 2020年12月10日
专知会员服务
41+阅读 · 2020年12月1日
相关基金
国家自然科学基金
0+阅读 · 2017年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员