AI agents can now take irreversible actions in operational systems, but agent-caused losses are still not clearly assigned, priced, or transferred. Providers often disclaim consequential damages, users are left with uncompensated losses, and default human review limits the efficiency gains of automation. We ask when autonomous AI deployment can become economically acceptable despite failure risk. Our answer is to quantify risk at the customer-task-trace episode level and transfer it through insurance. Automation is acceptable when its expected benefit exceeds the premium, control cost, and remaining risk. This requires a defined role with bounded permissions and comparable traces. We introduce trace-economic underwriting, which maps tool-use traces to customer exposure and claimable loss, then uses this representation for pricing, control, and risk transfer. It uses deterministic economic labels rather than an LLM judge. In our trace-to-loss testbed, trace-economic pricing reduces pricing MAE from $17.7K to $569 and removes regressive cross-subsidy. A 300-trace expert audit accepts 295 labels unchanged. On 1,000 real SWE-smith traces, trace-conditioned controls reduce CVaR95 by 72%. Theorem~1 gives a finite-sample scope condition. We release code, labels, and audit sheets.


翻译:AI代理如今能够在运营系统中执行不可逆的操作,但由此导致的损失仍未得到明确分配、定价或转移。提供商往往免除间接损失责任,用户承担未补偿的损失,而默认的人工审核则限制了自动化的效率提升。我们探讨的是:尽管存在失败风险,自主AI部署何时能变得经济上可接受。我们的解决方案是在客户-任务-轨迹环节层面量化风险,并通过保险转移风险。当自动化带来的预期收益超过保费、控制成本及剩余风险时,该部署即可被接受。这需要定义明确的角色、有限的权限以及可比较的轨迹。我们提出轨迹经济承保方法,将工具使用轨迹映射至客户敞口与可索赔损失,进而利用该表示进行定价、控制与风险转移。该方法采用确定性经济标签而非大语言模型评判器。在我们的轨迹到损失测试平台中,轨迹经济定价将定价平均绝对误差从17700美元降至569美元,并消除了逆向交叉补贴。一项基于300条轨迹的专家审计接受了其中295个标签未经修改。在1000条真实SWE-smith轨迹上,轨迹条件化控制将CVaR95降低了72%。定理1给出了有限样本范围条件。我们已发布代码、标签及审计表格。

0
下载
关闭预览

相关内容

自主智能:多模态人工智能代理重塑技术未来
专知会员服务
26+阅读 · 2025年11月23日
基于动态知识图谱的人工智能代理自主研究周期 | 文献
专知会员服务
27+阅读 · 2025年10月24日
一种Agent自主性风险评估框架 | 最新文献
专知会员服务
24+阅读 · 2025年10月24日
AI Agent:基于大模型的自主智能体
专知会员服务
251+阅读 · 2023年9月9日
人工智能商业化研究报告(2019)
腾讯大讲堂
15+阅读 · 2019年7月9日
自动驾驶汽车技术路线简介
智能交通技术
15+阅读 · 2019年4月25日
人工智能对网络空间安全的影响
走向智能论坛
21+阅读 · 2018年6月7日
【人工智能】人工智能5大商业模式
产业智能官
17+阅读 · 2017年10月16日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
4+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
12+阅读 · 2013年12月31日
VIP会员
最新内容
从采集到决策:美军视角下的战术情报范式重构
专知会员服务
1+阅读 · 今天2:42
《履带式无人地面战车技术发展现状》
专知会员服务
2+阅读 · 今天1:46
《无人机脆弱性利用:网络空间力量的新域》
专知会员服务
2+阅读 · 8月1日
美空军如何将人工智能从战场部署至后方机关
专知会员服务
11+阅读 · 7月31日
《史诗怒火行动:多域前瞻评估》49页报告
专知会员服务
7+阅读 · 7月31日
《英国防部:未来空战系统数字化战略》33页
专知会员服务
5+阅读 · 7月31日
《面向自主飞行网络的智能体人工智能架构》
专知会员服务
7+阅读 · 7月31日
相关基金
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
4+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
12+阅读 · 2013年12月31日
Top
微信扫码咨询专知VIP会员