Search intelligence is evolving from Deep Research to Wide Research, a paradigm essential for retrieving and synthesizing comprehensive information under complex constraints in parallel. However, progress in this field is impeded by the lack of dedicated benchmarks and optimization methodologies for search breadth. To address these challenges, we take a deep dive into Wide Research from two perspectives: Data Pipeline and Agent Optimization. First, we produce WideSeekBench, a General Broad Information Seeking (GBIS) benchmark constructed via a rigorous multi-phase data pipeline to ensure diversity across the target information volume, logical constraints, and domains. Second, we introduce WideSeek, a dynamic hierarchical multi-agent architecture that can autonomously fork parallel sub-agents based on task requirements. Furthermore, we design a unified training framework that linearizes multi-agent trajectories and optimizes the system using end-to-end RL. Experimental results demonstrate the effectiveness of WideSeek and multi-agent RL, highlighting that scaling the number of agents is a promising direction for advancing the Wide Research paradigm.


翻译:搜索智能正从深度研究向广度研究演进,后者是一种在复杂约束下并行检索与综合全面信息的关键范式。然而,该领域的发展因缺乏针对搜索广度的专用基准与优化方法而受阻。为应对这些挑战,我们从数据管道与智能体优化两个视角深入探究广度研究。首先,我们构建了WideSeekBench,这是一个通过严格多阶段数据管道构建的通用广域信息寻求基准,旨在确保目标信息量、逻辑约束与领域多样性。其次,我们提出了WideSeek,一种动态分层多智能体架构,能够根据任务需求自主分叉并行子智能体。此外,我们设计了一个统一的训练框架,将多智能体轨迹线性化,并利用端到端强化学习对系统进行优化。实验结果验证了WideSeek与多智能体强化学习的有效性,表明扩展智能体数量是推进广度研究范式的可行方向。

0
下载
关闭预览

相关内容

智能体,顾名思义,就是具有智能的实体,英文名是Agent。
【NUS博士论文】面向交互的多智能体行为预测,156页pdf
专知会员服务
32+阅读 · 2024年11月17日
多智能体深度强化学习研究进展
专知会员服务
76+阅读 · 2024年7月17日
多智能体博弈学习研究进展
专知会员服务
91+阅读 · 2024年5月5日
基于学习机制的多智能体强化学习综述
专知会员服务
64+阅读 · 2024年4月16日
专知会员服务
172+阅读 · 2021年8月3日
专知会员服务
214+阅读 · 2019年8月30日
【综述】多智能体强化学习算法理论研究
深度强化学习实验室
16+阅读 · 2020年9月9日
【DeepMind】多智能体学习231页PPT总结
深度强化学习实验室
16+阅读 · 2020年6月23日
浅谈群体智能——新一代AI的重要方向
中国科学院自动化研究所
44+阅读 · 2019年10月16日
群体智能:新一代人工智能的重要方向
走向智能论坛
12+阅读 · 2017年8月16日
国家自然科学基金
4+阅读 · 2017年12月31日
国家自然科学基金
43+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
13+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
10+阅读 · 2013年12月31日
国家自然科学基金
18+阅读 · 2009年12月31日
国家自然科学基金
50+阅读 · 2009年12月31日
Arxiv
0+阅读 · 2月27日
VIP会员
最新内容
美空军新型反无人机部队初探
专知会员服务
1+阅读 · 今天5:45
《防空交战流程的概率建模研究》
专知会员服务
4+阅读 · 今天5:04
ICML 2026 教程 | 数值优化理论还重要吗?
专知会员服务
4+阅读 · 7月26日
ICM 2026 | 陶哲轩:人工智能时代的数学
专知会员服务
7+阅读 · 7月26日
《反无人机交战场景下的战斗归零研究》
专知会员服务
7+阅读 · 7月26日
博士论文 | 用代码结构感知方法推进代码大模型
相关VIP内容
【NUS博士论文】面向交互的多智能体行为预测,156页pdf
专知会员服务
32+阅读 · 2024年11月17日
多智能体深度强化学习研究进展
专知会员服务
76+阅读 · 2024年7月17日
多智能体博弈学习研究进展
专知会员服务
91+阅读 · 2024年5月5日
基于学习机制的多智能体强化学习综述
专知会员服务
64+阅读 · 2024年4月16日
专知会员服务
172+阅读 · 2021年8月3日
专知会员服务
214+阅读 · 2019年8月30日
相关基金
国家自然科学基金
4+阅读 · 2017年12月31日
国家自然科学基金
43+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
13+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
10+阅读 · 2013年12月31日
国家自然科学基金
18+阅读 · 2009年12月31日
国家自然科学基金
50+阅读 · 2009年12月31日
Top
微信扫码咨询专知VIP会员