Limited work has examined the strategic behaviors of relational networked learning agents under social dilemmas, and has overlooked the intricate social dynamics of complex systems. We address the challenge with Socio-Relational Intrinsic Motivation (SRIM), which endows agents with diverse preferences over sub-graphical social structures in order to study the impact of agents' personal preferences over their sub-graphical relations on their strategic decision-making under sequential social dilemmas. Our results in the Harvest and Cleanup environments demonstrate that preferences over different subgraph structures (degree-, clique-, and critical connection-based) lead to distinct variations in agents' reward gathering and strategic behavior: individual aggressiveness in Harvest and individual contribution effort in Cleanup. Moreover, agents with different subgraphical structural positions consistently exhibit similar strategic behavioral shifts. Our proposed BCI metric captures structural variation within the population, and the relative ordering of BCI across social preferences is consistent in Harvest and Cleanup games for the same topology, suggesting the subgraphical structural impact is robust across environments. These results provide a new lens for examining agents' behavior in social dilemmas and insight for designing effective multi-agent ecosystems composed of heterogeneous social agents.


翻译:关于关系型网络学习主体在社会困境下的策略行为研究有限,且忽视了复杂系统内精细的社会动力学。我们通过社会关系内在动机(SRIM)应对这一挑战,该机制赋予主体对子图社会结构的多样化偏好,旨在研究主体对子图关系的个人偏好如何影响其在序贯社会困境下的策略决策。我们在Harvest和Cleanup环境中的实验表明:对不同子图结构(基于度数、团簇和关键连接)的偏好会导致主体奖励获取与策略行为的显著差异——Harvest中的个体攻击性及Cleanup中的个体贡献努力程度。此外,处于不同子图结构位置的主体始终表现出相似的策略行为转变。我们提出的BCI指标能够捕获种群内的结构变异,且相同拓扑结构下不同社会偏好的BCI相对排序在Harvest与Cleanup游戏中具有一致性,表明子图结构影响具有跨环境稳健性。这些成果为审视社会困境中的主体行为提供了新视角,并为设计由异构社会主体构成的有效多智能体生态系统提供了洞见。

0
下载
关闭预览

相关内容

《分布式多智能体强化学习策略的可解释性研究》
专知会员服务
30+阅读 · 2025年11月17日
直接偏好优化中的数据集、理论、变体和应用的综合综述
专知会员服务
15+阅读 · 2024年10月24日
佐治亚理工学院最新《图神经网络社会推荐系统》2022综述
图机器学习 2.2-2.4 Properties of Networks, Random Graph
图与推荐
10+阅读 · 2020年3月28日
网络表示学习概述
机器学习与推荐算法
20+阅读 · 2020年3月27日
ACL 2019开源论文 | 基于Attention的知识图谱关系预测
用深度学习揭示数据的因果关系
专知
28+阅读 · 2019年5月18日
基于注意力机制的图卷积网络
科技创新与创业
74+阅读 · 2017年11月8日
特定目标情感分析——神经网络这是要逆天么
计算机研究与发展
14+阅读 · 2017年9月5日
国家自然科学基金
1+阅读 · 2016年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
6+阅读 · 2014年12月31日
VIP会员
最新内容
俄乌无人机战争的六大启示
专知会员服务
9+阅读 · 8月3日
《无人机空中监控:通信实验洞察》
专知会员服务
6+阅读 · 8月3日
从采集到决策:美军视角下的战术情报范式重构
《履带式无人地面战车技术发展现状》
专知会员服务
6+阅读 · 8月2日
《无人机脆弱性利用:网络空间力量的新域》
专知会员服务
9+阅读 · 8月1日
美空军如何将人工智能从战场部署至后方机关
专知会员服务
14+阅读 · 7月31日
相关VIP内容
《分布式多智能体强化学习策略的可解释性研究》
专知会员服务
30+阅读 · 2025年11月17日
直接偏好优化中的数据集、理论、变体和应用的综合综述
专知会员服务
15+阅读 · 2024年10月24日
佐治亚理工学院最新《图神经网络社会推荐系统》2022综述
相关资讯
图机器学习 2.2-2.4 Properties of Networks, Random Graph
图与推荐
10+阅读 · 2020年3月28日
网络表示学习概述
机器学习与推荐算法
20+阅读 · 2020年3月27日
ACL 2019开源论文 | 基于Attention的知识图谱关系预测
用深度学习揭示数据的因果关系
专知
28+阅读 · 2019年5月18日
基于注意力机制的图卷积网络
科技创新与创业
74+阅读 · 2017年11月8日
特定目标情感分析——神经网络这是要逆天么
计算机研究与发展
14+阅读 · 2017年9月5日
相关基金
国家自然科学基金
1+阅读 · 2016年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
6+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员