成为VIP会员查看完整内容
VIP会员码认证
首页
主题
会员
服务
注册
·
登录
大语言模型智能体
关注
5
综合
百科
VIP
热门
动态
论文
精华
Consensus-based Agentic Large Language Model Framework for Harmonized Tariff Schedule Code Classification
Arxiv
0+阅读 · 6月15日
From Agent Traces to Trust: A Survey of Evidence Tracing and Execution Provenance in LLM Agents
Arxiv
0+阅读 · 6月14日
How Many Tools Should an LLM Agent See? A Chance-Corrected Answer
Arxiv
0+阅读 · 6月7日
Retrospective Progress-Aware Self-Refinement for LLM Agent Training
Arxiv
0+阅读 · 6月12日
Voluntary Collusion with Secret Tools in Competing LLM Agents
Arxiv
0+阅读 · 5月26日
More Skills, Worse Agents? Skill Shadowing Degrades Performance When Expanding Skill Libraries
Arxiv
0+阅读 · 5月21日
SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories
Arxiv
0+阅读 · 5月31日
An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations
Arxiv
0+阅读 · 5月28日
From Failed Trajectories to Reliable LLM Agents: Diagnosing and Repairing Harness Flaws
Arxiv
0+阅读 · 6月4日
Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems
Arxiv
0+阅读 · 5月27日
Layer-Isolated Evaluation: Gating the Deterministic Scaffold of a Production LLM Agent with a No-LLM, Regression-Locked Test Harness
Arxiv
0+阅读 · 6月10日
AgentAtlas: Beyond Outcome Leaderboards for LLM Agents
Arxiv
0+阅读 · 5月26日
Infini Memory: Maintainable Topic Documents for Long-Term LLM Agent Memory
Arxiv
0+阅读 · 6月9日
Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense
Arxiv
0+阅读 · 6月4日
LLM-Agent-based Social Simulation for Attitude Diffusion
Arxiv
0+阅读 · 4月4日
参考链接
提示
微信扫码
咨询专知VIP会员与技术项目合作
(加微信请备注: "专知")
微信扫码咨询专知VIP会员
Top