AI governance frameworks increasingly emphasize fairness, transparency, accountability, and lifecycle risk management in high-stakes domains. However, many current approaches remain observational, relying on static metric reporting, post-hoc auditing, and monitoring dashboards without directly governing deployment readiness, remediation progression, escalation states, or assurance-driven deployment control. This paper introduces Operational AI Deployment Assurance (OADA), a governance framework for translating fairness disagreement, subgroup instability, threshold sensitivity, remediation outcomes, and operational uncertainty into deployment-oriented assurance decisions. Building on prior work on the Fairness Disagreement Index (FDI) and FairRisk-FDI, OADA reframes governance uncertainty as an operational concern within AI deployment pipelines rather than a byproduct of metric disagreement. The framework introduces Deployment Assurance Scores, Deployment Readiness Classifications, Threshold Stability Zones, Governance Escalation States, and remediation-aware assurance progression. These constructs support lifecycle-oriented governance decisions across high-stakes settings by connecting evaluation outputs to deployment-state interpretation, reassessment, escalation, and operational control. Through deployment-oriented evaluation across facial recognition systems, with discussion extended to healthcare AI as a representative high-stakes domain, the paper demonstrates how systems may appear acceptable under isolated fairness or performance metrics while still exhibiting instability that affects deployment readiness. The proposed framework positions operational deployment assurance as a governance layer between evaluation and real-world AI deployment.


翻译:人工智能治理框架日益强调高风险领域中的公平性、透明度、问责制和全生命周期风险管理。然而,当前许多方法仍停留在观察层面,依赖静态指标报告、事后审计和监控仪表盘,未能直接管理部署就绪度、修复进程、升级状态或基于保障的部署控制。本文提出操作化部署保障(OADA),这是一种将公平性分歧、子群不稳定性、阈值敏感性、修复结果及操作不确定性转化为面向部署的保障决策的治理框架。基于先前关于公平性分歧指数(FDI)和FairRisk-FDI的研究,OADA将治理不确定性重新定义为AI部署流水线中的操作性问题,而非指标分歧的副产品。该框架引入了部署保障评分、部署就绪度分类、阈值稳定区间、治理升级状态以及考虑修复的保障进程。这些构造通过将评估输出与部署状态解读、重新评估、升级和操作控制相连接,支持跨高风险场景的生命周期治理决策。通过在面部识别系统中开展面向部署的评估,并将讨论延伸至代表性高风险领域(医疗AI),本文展示了系统在孤立的公平性或性能指标下可能看似可接受,但仍存在影响部署就绪度的不稳定性。所提出的框架将操作化部署保障定位为评估与现实世界AI部署之间的治理层。

0
下载
关闭预览

相关内容

《人工智能使能系统可靠性框架》
专知会员服务
20+阅读 · 4月27日
《军用AI智能体的治理框架》最新报告
专知会员服务
38+阅读 · 3月8日
一种Agent自主性风险评估框架 | 最新文献
专知会员服务
24+阅读 · 2025年10月24日
《人工智能军事系统的风险分级监管路径》
专知会员服务
23+阅读 · 2025年7月10日
中国信通院发布《人工智能风险治理报告(2024年)》
专知会员服务
49+阅读 · 2024年12月26日
新加坡-生成式AI的治理框架模型,23页pdf
专知会员服务
59+阅读 · 2024年2月4日
专知会员服务
65+阅读 · 2021年7月5日
《人工智能安全框架(2020年)》白皮书,68页pdf
专知会员服务
167+阅读 · 2021年1月9日
重磅!AI框架发展白皮书(2022年),44页pdf
专知
28+阅读 · 2022年2月27日
《人工智能安全测评白皮书》,99页pdf
专知
36+阅读 · 2022年2月26日
智能时代如何构建金融反欺诈体系?
数据猿
12+阅读 · 2018年3月26日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
18+阅读 · 2014年12月31日
国家自然科学基金
18+阅读 · 2012年12月31日
国家自然科学基金
24+阅读 · 2011年12月31日
VIP会员
最新内容
非对称防御中的自组织临界性:俄乌战争
专知会员服务
7+阅读 · 8月10日
《战争中的大语言模型监管》
专知会员服务
6+阅读 · 8月10日
《边缘计算关键技术分析及美军作战实践应用》
边缘计算的军事应用
专知会员服务
11+阅读 · 8月9日
一种考虑资源机动性的武器目标分配混合算法
专知会员服务
12+阅读 · 8月8日
相关VIP内容
《人工智能使能系统可靠性框架》
专知会员服务
20+阅读 · 4月27日
《军用AI智能体的治理框架》最新报告
专知会员服务
38+阅读 · 3月8日
一种Agent自主性风险评估框架 | 最新文献
专知会员服务
24+阅读 · 2025年10月24日
《人工智能军事系统的风险分级监管路径》
专知会员服务
23+阅读 · 2025年7月10日
中国信通院发布《人工智能风险治理报告(2024年)》
专知会员服务
49+阅读 · 2024年12月26日
新加坡-生成式AI的治理框架模型,23页pdf
专知会员服务
59+阅读 · 2024年2月4日
专知会员服务
65+阅读 · 2021年7月5日
《人工智能安全框架(2020年)》白皮书,68页pdf
专知会员服务
167+阅读 · 2021年1月9日
相关资讯
重磅!AI框架发展白皮书(2022年),44页pdf
专知
28+阅读 · 2022年2月27日
《人工智能安全测评白皮书》,99页pdf
专知
36+阅读 · 2022年2月26日
智能时代如何构建金融反欺诈体系?
数据猿
12+阅读 · 2018年3月26日
相关基金
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
18+阅读 · 2014年12月31日
国家自然科学基金
18+阅读 · 2012年12月31日
国家自然科学基金
24+阅读 · 2011年12月31日
Top
微信扫码咨询专知VIP会员