Gödel agent realize recursive self-improvement: an agent inspects its own policy and traces and then modifies that policy in a tested loop. We introduce Polaris, a Gödel agent for compact models that performs policy repair via experience abstraction, turning failures into policy updates through a structured cycle of analysis, strategy formation, abstraction, and minimal code pat ch repair with conservative checks. Unlike response level self correction or parameter tuning, Polaris makes policy level changes with small, auditable patches that persist in the policy and are reused on unseen instances within each benchmark. As part of the loop, the agent engages in meta reasoning: it explains its errors, proposes concrete revisions to its own policy, and then updates the policy. To enable cumulative policy refinement, we introduce experience abstraction, which distills failures into compact, reusable strategies that transfer to unseen instances. On MGSM, DROP, GPQA, and LitBench (covering arithmetic reasoning, compositional inference, graduate-level problem solving, and creative writing evaluation), a 7-billion-parameter model equipped with Polaris achieves consistent gains over the base policy and competitive baselines.


翻译:摘要:哥德尔智能体(Gödel agent)实现递归式自我改进:智能体检查其自身策略与轨迹,并在经过验证的循环中修改该策略。我们提出 Polaris——一种面向紧凑模型的哥德尔智能体,通过经验抽象进行策略修复,将失败转化为策略更新。该过程采用结构化循环,涵盖分析、策略形成、抽象、最小化代码补丁修复及保守性检查。与响应级自我纠正或参数调优不同,Polaris 通过小型可审计补丁实现策略级更改,这些补丁持久存在于策略中,并在每个基准测试中用于未见实例。在该循环中,智能体进行元推理:解释自身错误,提出针对其策略的具体修订方案,并最终更新策略。为支持累积性策略优化,我们引入经验抽象机制,将失败案例提炼为可迁移至未见实例的紧凑可重用策略。在涵盖算术推理、组合推理、研究生级问题求解与创意写作评估的 MGSM、DROP、GPQA 及 LitBench 基准测试中,配备 Polaris 的 70 亿参数模型较基础策略及竞争基线均取得持续增益。

0
下载
关闭预览

相关内容

智能体,顾名思义,就是具有智能的实体,英文名是Agent。
《基于Transformer的智能体的战术决策解释》
专知会员服务
49+阅读 · 2025年12月28日
大语言模型智能体强化学习:全景综述
专知会员服务
51+阅读 · 2025年12月18日
智能体化多模态大语言模型综述
专知会员服务
40+阅读 · 2025年10月14日
面向大语言模型的智能体化强化学习图景:综述
专知会员服务
56+阅读 · 2025年9月3日
基于大语言模型的智能体优化研究综述
专知会员服务
65+阅读 · 2025年3月25日
超越ChatGPT的AI智能体,82页ppt
专知会员服务
56+阅读 · 2025年2月15日
基于大型语言模型的软件工程智能体综述
专知会员服务
61+阅读 · 2024年9月6日
强化学习《奖励函数设计: Reward Shaping》详细解读
深度强化学习实验室
20+阅读 · 2020年9月1日
PlaNet 简介:用于强化学习的深度规划网络
谷歌开发者
13+阅读 · 2019年3月16日
基于python的开源量化交易,量化投资架构
运维帮
15+阅读 · 2018年7月5日
TextInfoExp:自然语言处理相关实验(基于sougou数据集)
全球人工智能
12+阅读 · 2017年11月12日
群体智能:新一代人工智能的重要方向
走向智能论坛
12+阅读 · 2017年8月16日
国家自然科学基金
43+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
47+阅读 · 2015年12月31日
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
10+阅读 · 2013年12月31日
国家自然科学基金
18+阅读 · 2009年12月31日
国家自然科学基金
17+阅读 · 2008年12月31日
Arxiv
14+阅读 · 2023年8月7日
VIP会员
最新内容
乌克兰纵深打击如何重塑俄罗斯的战略选择
专知会员服务
1+阅读 · 今天12:25
俄乌战争中关于中程打击无人机部署的经验启示
专知会员服务
0+阅读 · 今天12:08
《基于强化学习的自动化红队测试》
专知会员服务
4+阅读 · 7月23日
伊朗不对称防空战略的演进
专知会员服务
4+阅读 · 7月23日
对抗环境下超视距目标打击的情报支援
专知会员服务
10+阅读 · 7月22日
相关VIP内容
《基于Transformer的智能体的战术决策解释》
专知会员服务
49+阅读 · 2025年12月28日
大语言模型智能体强化学习:全景综述
专知会员服务
51+阅读 · 2025年12月18日
智能体化多模态大语言模型综述
专知会员服务
40+阅读 · 2025年10月14日
面向大语言模型的智能体化强化学习图景:综述
专知会员服务
56+阅读 · 2025年9月3日
基于大语言模型的智能体优化研究综述
专知会员服务
65+阅读 · 2025年3月25日
超越ChatGPT的AI智能体,82页ppt
专知会员服务
56+阅读 · 2025年2月15日
基于大型语言模型的软件工程智能体综述
专知会员服务
61+阅读 · 2024年9月6日
相关基金
国家自然科学基金
43+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
47+阅读 · 2015年12月31日
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
10+阅读 · 2013年12月31日
国家自然科学基金
18+阅读 · 2009年12月31日
国家自然科学基金
17+阅读 · 2008年12月31日
Top
微信扫码咨询专知VIP会员