In the realm of machine learning (ML) systems featuring client-host connections, the enhancement of privacy security can be effectively achieved through federated learning (FL) as a secure distributed ML methodology. FL effectively integrates cloud infrastructure to transfer ML models onto edge servers using blockchain technology. Through this mechanism, it guarantees the streamlined processing and data storage requirements of both centralized and decentralized systems, with an emphasis on scalability, privacy considerations, and cost-effective communication. In current FL implementations, data owners locally train their models, and subsequently upload the outcomes in the form of weights, gradients, and parameters to the cloud for overall model aggregation. This innovation obviates the necessity of engaging Internet of Things (IoT) clients and participants to communicate raw and potentially confidential data directly with a cloud center. This not only reduces the costs associated with communication networks but also enhances the protection of private data. This survey conducts an analysis and comparison of recent FL applications, aiming to assess their efficiency, accuracy, and privacy protection. However, in light of the complex and evolving nature of FL, it becomes evident that additional research is imperative to address lingering knowledge gaps and effectively confront the forthcoming challenges in this field. In this study, we categorize recent literature into the following clusters: privacy protection, resource allocation, case study analysis, and applications. Furthermore, at the end of each section, we tabulate the open areas and future directions presented in the referenced literature, affording researchers and scholars an insightful view of the evolution of the field.
翻译:在采用客户端-服务器连接的机器学习系统中,联邦学习作为一种安全的分布式机器学习方法,能够有效增强隐私安全性。联邦学习通过利用区块链技术,将云基础设施整合到边缘服务器上传输机器学习模型。通过这种机制,它保证了集中式和分布式系统在可扩展性、隐私考量和经济高效通信方面的流线化处理与数据存储需求。在当前的联邦学习实现中,数据所有者本地训练其模型,随后将权重、梯度和参数形式的训练结果上传至云端进行全局模型聚合。这一创新消除了物联网客户端和参与者需要直接与云中心通信原始且可能敏感数据的必要性,不仅降低了通信网络的成本,还增强了私有数据的保护。本综述对近期联邦学习的应用进行了分析比较,旨在评估其效率、准确性和隐私保护能力。然而,鉴于联邦学习的复杂性和不断演变的特性,显然需要开展更多研究来弥补现有知识空白,并有效应对该领域即将面临的挑战。在本研究中,我们将近期文献分为以下几类:隐私保护、资源分配、案例研究分析和应用。此外,在每节末尾,我们以表格形式列出了参考文献中提出的开放领域和未来方向,为研究人员和学者提供了对该领域演变的深入洞察。