Shor's algorithm proved that asymmetric cryptographic protocols based on the integer factorization and discrete logarithm problems are no longer safe in a world with large-scale quantum computers. As a result, Post-Quantum Cryptography (PQC) has been developed over the last few years, seeking cryptographic primitives resistant to quantum attacks. One of the main hard problems underlying PQC schemes is the Learning with Errors (LWE) problem, which is significantly more computationally intensive than its classical predecessors. In this work, we present a Key Encapsulation Mechanism (KEM) based on plain LWE and develop a GPU-oriented implementation using OpenACC. We evaluate the performance of our accelerated application in terms of both time-to-solution and energy-to-solution, considering bare-metal and containerized executions across multiple NVIDIA GPU models and generations. Our implementation achieves significant acceleration across all tested GPU platforms. In particular, on the NVIDIA Grace Hopper Superchip, it attains up to a $208\times$ speedup over a multithreaded CPU baseline and enables the execution of problem sizes that are impractical on CPU architectures due to memory and synchronization constraints. Energy consumption analysis also shows $\approx 2\times$ better efficiency when using the Superchip compared to systems equipped with x86-based CPUs and NVIDIA H100 GPUs. These results highlight the effectiveness of GPU acceleration for computationally demanding LWE-based cryptographic workloads.


翻译:Shor算法证明了在拥有大规模量子计算机的世界中,基于整数分解和离散对数问题的非对称密码协议将不再安全。因此,在过去几年中,后量子密码学(Post-Quantum Cryptography, PQC)得到了发展,旨在寻找能够抵抗量子攻击的密码学原语。PQC方案所依赖的主要困难问题之一是错误学习问题(Learning with Errors, LWE),其计算强度远高于经典的前身。本文提出了一种基于原始LWE的密钥封装机制(Key Encapsulation Mechanism, KEM),并利用OpenACC开发了面向GPU的实现。我们从求解时间和能耗两个角度评估了加速应用的性能,考虑了在多种NVIDIA GPU模型和代际上进行裸机和容器化执行的情况。我们的实现在所有测试的GPU平台上均实现了显著的加速。特别是,在NVIDIA Grace Hopper Superchip上,相较于多线程CPU基线,实现了高达208倍的加速比,并能够执行在CPU架构上因内存和同步约束而不可行的问题规模。能耗分析还表明,与配备x86 CPU和NVIDIA H100 GPU的系统相比,使用Superchip的效率提高了约2倍。这些结果突显了GPU加速对计算密集型的基于LWE的密码工作负载的有效性。

0
下载
关闭预览

相关内容

《量子机器学习》最新综述
专知会员服务
41+阅读 · 2023年8月24日
【2022新书】给工程师的量子机器学习简介,204页pdf
专知会员服务
56+阅读 · 2022年5月22日
专知会员服务
39+阅读 · 2021年10月19日
专知会员服务
25+阅读 · 2020年9月14日
用深度学习揭示数据的因果关系
专知
28+阅读 · 2019年5月18日
机器学习的Pytorch实现资源集合
专知
11+阅读 · 2018年9月1日
荐书丨OpenCV算法精解:基于Python与C++
程序人生
19+阅读 · 2017年11月18日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
VIP会员
最新内容
从采集到决策:美军视角下的战术情报范式重构
专知会员服务
1+阅读 · 今天2:42
《履带式无人地面战车技术发展现状》
专知会员服务
2+阅读 · 今天1:46
《无人机脆弱性利用:网络空间力量的新域》
专知会员服务
2+阅读 · 8月1日
美空军如何将人工智能从战场部署至后方机关
专知会员服务
11+阅读 · 7月31日
《史诗怒火行动:多域前瞻评估》49页报告
专知会员服务
7+阅读 · 7月31日
《英国防部:未来空战系统数字化战略》33页
专知会员服务
5+阅读 · 7月31日
《面向自主飞行网络的智能体人工智能架构》
专知会员服务
7+阅读 · 7月31日
相关基金
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员