We identify a fundamental incompatibility between the goals of accuracy, trust, and human-level reasoning in artificial intelligence (AI) systems, for strict mathematical definitions of these notions. We define accuracy of a system as the property that it never makes any false claims when it has the ability to abstain from making a prediction on any input, and trust as the assumption that the system is accurate. We define human-level reasoning as the property of an AI system always matching or exceeding human capability. Our core finding is that -- for our formal definitions of these notions -- an accurate and trusted AI system cannot be a human-level reasoning system: for such an accurate, trusted system there are task instances which are easily and provably solvable by a human but not by the system. Our proofs draw parallels to Gödel's incompleteness theorems and Turing's proof of the undecidability of the halting problem, and can be regarded as interpretations of Gödel's and Turing's results. Key to our proof is the formalization of the notion of trust, which allows us to separate the intrinsic property of a system (being accurate) from its epistemic status (being trusted).


翻译:我们发现在严格数学定义下,人工智能系统中的精确性、可信性与人类水平推理目标之间存在根本性矛盾。我们将系统的精确性定义为:当其具备对任意输入可放弃预测的能力时,系统不会作出任何错误断言的性质;将可信性定义为系统精确性的预设前提。我们将人类水平推理定义为人工智能系统始终达到或超越人类能力水平的性质。核心发现是:基于对这些概念的正式定义,一个精确且可信的人工智能系统无法成为人类水平推理系统——对于这类精确可信的系统,存在人类可以轻松、可证明地解决但系统无法解决的任务实例。我们的证明与哥德尔不完备定理及图灵对停机问题不可判定性的证明存在相似性,可视为对哥德尔与图灵定理的诠释。证明的关键在于对可信性概念的形式化,这使我们得以区分系统的内在属性(精确性)与其认知地位(可信性)。

0
下载
关闭预览

相关内容

可解释人工智能的基础
专知会员服务
33+阅读 · 2025年10月26日
【CMU博士论文】基于机器学习的可信科学推理
专知会员服务
17+阅读 · 2025年5月26日
《提高决策支持系统透明度的可解释人工智能》最新100页
专知会员服务
52+阅读 · 2024年11月28日
深度学习在数学推理中的应用综述
专知会员服务
49+阅读 · 2022年12月25日
人工智能系统可信性度量评估研究综述
专知会员服务
99+阅读 · 2022年1月30日
【机器推理可解释性】Machine Reasoning Explainability
专知会员服务
35+阅读 · 2020年9月3日
「因果推理」概述论文,13页pdf
专知
16+阅读 · 2021年3月20日
最新《可解释人工智能》概述,50页ppt
专知
12+阅读 · 2021年3月17日
理解人类推理的深度学习
论智
19+阅读 · 2018年11月7日
【混合智能】人机混合智能的哲学思考
产业智能官
12+阅读 · 2018年10月28日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
18+阅读 · 2012年12月31日
国家自然科学基金
23+阅读 · 2008年12月31日
Arxiv
0+阅读 · 5月31日
VIP会员
最新内容
《履带式无人地面战车技术发展现状》
专知会员服务
2+阅读 · 8月2日
《无人机脆弱性利用:网络空间力量的新域》
专知会员服务
2+阅读 · 8月1日
美空军如何将人工智能从战场部署至后方机关
专知会员服务
11+阅读 · 7月31日
《史诗怒火行动:多域前瞻评估》49页报告
专知会员服务
8+阅读 · 7月31日
《英国防部:未来空战系统数字化战略》33页
专知会员服务
5+阅读 · 7月31日
《面向自主飞行网络的智能体人工智能架构》
专知会员服务
8+阅读 · 7月31日
相关基金
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
18+阅读 · 2012年12月31日
国家自然科学基金
23+阅读 · 2008年12月31日
Top
微信扫码咨询专知VIP会员