The dangers of adversarial attacks on Uncrewed Aerial Vehicle (UAV) agents operating in public are increasing. Adopting AI-based techniques and, more specifically, Deep Learning (DL) approaches to control and guide these UAVs can be beneficial in terms of performance but can add concerns regarding the safety of those techniques and their vulnerability against adversarial attacks. Confusion in the agent's decision-making process caused by these attacks can seriously affect the safety of the UAV. This paper proposes an innovative approach based on the explainability of DL methods to build an efficient detector that will protect these DL schemes and the UAVs adopting them from attacks. The agent adopts a Deep Reinforcement Learning (DRL) scheme for guidance and planning. The agent is trained with a Deep Deterministic Policy Gradient (DDPG) with Prioritised Experience Replay (PER) DRL scheme that utilises Artificial Potential Field (APF) to improve training times and obstacle avoidance performance. A simulated environment for UAV explainable DRL-based planning and guidance, including obstacles and adversarial attacks, is built. The adversarial attacks are generated by the Basic Iterative Method (BIM) algorithm and reduced obstacle course completion rates from 97\% to 35\%. Two adversarial attack detectors are proposed to counter this reduction. The first one is a Convolutional Neural Network Adversarial Detector (CNN-AD), which achieves accuracy in the detection of 80\%. The second detector utilises a Long Short Term Memory (LSTM) network. It achieves an accuracy of 91\% with faster computing times compared to the CNN-AD, allowing for real-time adversarial detection.


翻译:无人机在公共环境中运行时面临日益增长的对抗攻击风险。采用基于人工智能的技术,特别是深度学习方法来控制和引导这些无人机,虽然能提升性能,但也引发了这些技术的安全性及其对抗攻击脆弱性的担忧。攻击导致智能体决策过程混乱,可能严重影响无人机安全。本文提出一种创新方法,基于深度学习方法的可解释性构建高效检测器,以保护深度学习方案及采用这些方案的无人机免受攻击。智能体采用深度强化学习方案进行导航与规划。该智能体使用基于优先经验回放的深度确定性策略梯度深度强化学习方案进行训练,并利用人工势场改进训练时间和避障性能。本文构建了包含障碍物和对抗攻击的无人机可解释深度强化学习规划与导航仿真环境。对抗攻击采用基本迭代法算法生成,使障碍物航线完成率从97%降至35%。为应对该性能下降,本文提出两种对抗攻击检测器。第一种是卷积神经网络对抗检测器,检测准确率达80%;第二种采用长短期记忆网络,检测准确率达91%,且计算速度比卷积神经网络对抗检测器更快,可实现实时对抗检测。

0
下载
关闭预览

相关内容

CMU 2022最新课程《可信赖人工智能自主性(TAIAT)》
专知会员服务
29+阅读 · 2022年4月30日
深度学习模型鲁棒性研究综述
专知会员服务
98+阅读 · 2022年1月23日
NeurlPS2022推荐系统论文集锦
机器学习与推荐算法
1+阅读 · 2022年9月26日
Multi-Task Learning的几篇综述文章
深度学习自然语言处理
15+阅读 · 2020年6月15日
“CVPR 2020 接受论文列表 1470篇论文都在这了
深度强化学习简介
专知
30+阅读 · 2018年12月3日
强化学习族谱
CreateAMind
26+阅读 · 2017年8月2日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2011年12月31日
国家自然科学基金
0+阅读 · 2011年12月31日
国家自然科学基金
8+阅读 · 2009年12月31日
国家自然科学基金
0+阅读 · 2009年12月31日
国家自然科学基金
1+阅读 · 2008年12月31日
Arxiv
0+阅读 · 2023年5月9日
Anomalous Instance Detection in Deep Learning: A Survey
VIP会员
最新内容
驱动军事决策变革的顶尖人工智能指挥系统
专知会员服务
6+阅读 · 8月11日
非对称防御中的自组织临界性:俄乌战争
专知会员服务
10+阅读 · 8月10日
《战争中的大语言模型监管》
专知会员服务
12+阅读 · 8月10日
《边缘计算关键技术分析及美军作战实践应用》
边缘计算的军事应用
专知会员服务
12+阅读 · 8月9日
一种考虑资源机动性的武器目标分配混合算法
专知会员服务
13+阅读 · 8月8日
相关VIP内容
CMU 2022最新课程《可信赖人工智能自主性(TAIAT)》
专知会员服务
29+阅读 · 2022年4月30日
深度学习模型鲁棒性研究综述
专知会员服务
98+阅读 · 2022年1月23日
相关基金
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2011年12月31日
国家自然科学基金
0+阅读 · 2011年12月31日
国家自然科学基金
8+阅读 · 2009年12月31日
国家自然科学基金
0+阅读 · 2009年12月31日
国家自然科学基金
1+阅读 · 2008年12月31日
Top
微信扫码咨询专知VIP会员