Understanding how humans collaborate and communicate in teams is essential for improving human-agent teaming and AI-assisted decision-making. However, relying solely on data from large-scale user studies is impractical due to logistical, ethical, and practical constraints, necessitating synthetic models of multiple diverse human behaviors. Recently, agents powered by Large Language Models (LLMs) have been shown to emulate human-like behavior in social settings. But, obtaining a large set of diverse behaviors requires manual effort in the form of designing prompts. On the other hand, Quality Diversity (QD) optimization has been shown to be capable of generating diverse Reinforcement Learning (RL) agent behavior. In this work, we combine QD optimization with LLM-powered agents to iteratively search for prompts that generate diverse team behavior in a long-horizon, multi-step collaborative environment. We first show, through a human-subjects experiment, that humans exhibit diverse coordination and communication behavior in this domain. We then present a series of experiments showing that our approach captures behaviors that are difficult to observe without large-scale data collection, and a follow-up user study to show that these generated behaviors are human-like. Our findings highlight the combination of QD and LLM-powered agents as an effective tool for studying teaming and communication strategies in multi-agent collaboration.


翻译:理解人类如何在团队中协作与通信对于提升人机协同及AI辅助决策至关重要。然而,仅依赖大规模用户研究数据因后勤、伦理及实践限制而不可行,亟需合成多类多样化人类行为的模型。近期研究表明,基于大语言模型的智能体能够在社交情境中模拟类人行为。但获取大量多样化行为需通过设计提示进行人工操作。另一方面,质量多样性优化已被证实在强化学习智能体行为生成中具有多样性。本研究将质量多样性优化与大语言模型驱动的智能体相结合,在长时域多步协作环境中迭代搜索生成多样化团队行为的提示。我们首先通过人类受试者实验证明,人类在此领域展现出多样化的协调与通信行为。随后通过系列实验表明,我们的方法能够捕捉到缺乏大规模数据收集时难以观测的行为,并开展后续用户研究证实这些生成的行为具有类人性。研究结果凸显了质量多样性优化与大语言模型智能体的结合,是研究多智能体协作中团队协作与通信策略的有效工具。

0
下载
关闭预览

相关内容

《异构人类团队的协作决策过程混合建模研究》
《多智能体大语言模型系统的可靠决策研究》
专知会员服务
41+阅读 · 2月2日
智能体化多模态大语言模型综述
专知会员服务
40+阅读 · 2025年10月14日
【EPFL博士论文】大型语言模型时代的协作式智能体
专知会员服务
36+阅读 · 2025年5月16日
多智能体协作机制:大语言模型综述
专知会员服务
86+阅读 · 2025年1月14日
大语言模型算法演进综述
专知会员服务
81+阅读 · 2024年5月30日
【ChatGPT系列报告】AI大语言模型的原理、演进及算力测算
专知会员服务
151+阅读 · 2023年4月26日
绝对干货!NLP预训练模型:从transformer到albert
新智元
14+阅读 · 2019年11月10日
浅谈群体智能——新一代AI的重要方向
中国科学院自动化研究所
44+阅读 · 2019年10月16日
常用的模型集成方法介绍:bagging、boosting 、stacking
群体智能:新一代人工智能的重要方向
走向智能论坛
12+阅读 · 2017年8月16日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2014年12月31日
国家自然科学基金
21+阅读 · 2013年12月31日
国家自然科学基金
19+阅读 · 2012年12月31日
国家自然科学基金
18+阅读 · 2009年12月31日
国家自然科学基金
50+阅读 · 2009年12月31日
VIP会员
最新内容
俄乌无人机战争的六大启示
专知会员服务
8+阅读 · 8月3日
《无人机空中监控:通信实验洞察》
专知会员服务
6+阅读 · 8月3日
从采集到决策:美军视角下的战术情报范式重构
《履带式无人地面战车技术发展现状》
专知会员服务
6+阅读 · 8月2日
《无人机脆弱性利用:网络空间力量的新域》
专知会员服务
6+阅读 · 8月1日
美空军如何将人工智能从战场部署至后方机关
专知会员服务
13+阅读 · 7月31日
相关VIP内容
《异构人类团队的协作决策过程混合建模研究》
《多智能体大语言模型系统的可靠决策研究》
专知会员服务
41+阅读 · 2月2日
智能体化多模态大语言模型综述
专知会员服务
40+阅读 · 2025年10月14日
【EPFL博士论文】大型语言模型时代的协作式智能体
专知会员服务
36+阅读 · 2025年5月16日
多智能体协作机制:大语言模型综述
专知会员服务
86+阅读 · 2025年1月14日
大语言模型算法演进综述
专知会员服务
81+阅读 · 2024年5月30日
【ChatGPT系列报告】AI大语言模型的原理、演进及算力测算
专知会员服务
151+阅读 · 2023年4月26日
相关基金
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2014年12月31日
国家自然科学基金
21+阅读 · 2013年12月31日
国家自然科学基金
19+阅读 · 2012年12月31日
国家自然科学基金
18+阅读 · 2009年12月31日
国家自然科学基金
50+阅读 · 2009年12月31日
Top
微信扫码咨询专知VIP会员