Purpose - To characterise and assess the quality of published research evaluating artificial intelligence (AI) methods for ovarian cancer diagnosis or prognosis using histopathology data. Methods - A search of 5 sources was conducted up to 01/12/2022. The inclusion criteria required that research evaluated AI on histopathology images for diagnostic or prognostic inferences in ovarian cancer, including tubo-ovarian and peritoneal tumours. Reviews and non-English language articles were excluded. The risk of bias was assessed for every included model using PROBAST. Results - A total of 1434 research articles were identified, of which 36 were eligible for inclusion. These studies reported 62 models of interest, including 35 classifiers, 14 survival prediction models, 7 segmentation models, and 6 regression models. Models were developed using 1-1375 slides from 1-664 ovarian cancer patients. A wide array of outcomes were predicted, including overall survival (9/62), histological subtypes (7/62), stain quantity (6/62) and malignancy (5/62). Older studies used traditional machine learning (ML) models with hand-crafted features, while newer studies typically employed deep learning (DL) to automatically learn features and predict the outcome(s) of interest. All models were found to be at high or unclear risk of bias overall. Research was frequently limited by insufficient reporting, small sample sizes, and insufficient validation. Conclusion - Limited research has been conducted and none of the associated models have been demonstrated to be ready for real-world implementation. Recommendations are provided addressing underlying biases and flaws in study design, which should help inform higher-quality reproducible future research. Key aspects include more transparent and comprehensive reporting, and improved performance evaluation using cross-validation and external validations.


翻译:目的 - 评估利用人工智能(AI)方法通过组织病理学数据进行卵巢癌诊断或预后预测的已发表研究,并对其质量进行评价。方法 - 截至2022年12月1日,对5个数据源进行了检索。纳入标准要求研究评估AI在组织病理学图像中进行卵巢癌(包括输卵管-卵巢及腹膜肿瘤)诊断或预后推断的应用。排除综述及非英语文献。使用PROBAST工具对每项纳入模型进行偏倚风险评估。结果 - 共识别出1434篇研究文献,其中36篇符合纳入标准。这些研究报告了62个目标模型,包括35个分类器、14个生存预测模型、7个分割模型和6个回归模型。模型基于1-1375张切片进行开发,涉及1-664例卵巢癌患者。预测结果涵盖总生存期(9/62)、组织学亚型(7/62)、染色强度(6/62)及恶性程度(5/62)等多个指标。早期研究采用基于手工特征的传统机器学习(ML)模型,而近期研究通常使用深度学习(DL)自动学习特征并预测目标结果。所有模型均被评估为整体高偏倚风险或偏倚风险不明确。研究普遍受限于报告不足、样本量小及验证不充分。结论 - 目前相关研究有限,且尚无模型被证明可实际应用于真实场景。针对研究设计中的潜在偏倚与缺陷提出了建议,以促进未来高质量、可重复的研究。关键方面包括更透明全面的报告机制,以及利用交叉验证和外部验证改进的性能评估方法。

0
下载
关闭预览

相关内容

Meta最新WWW2022《联邦计算导论》教程,附77页ppt
专知会员服务
60+阅读 · 2022年5月5日
【MIT干货课程】医疗健康领域的机器学习
专知
1+阅读 · 2022年5月26日
AI可解释性文献列表
专知
43+阅读 · 2019年10月7日
Hierarchically Structured Meta-learning
CreateAMind
27+阅读 · 2019年5月22日
A Technical Overview of AI & ML in 2018 & Trends for 2019
待字闺中
18+阅读 · 2018年12月24日
LibRec 精选:推荐系统的论文与源码
LibRec智能推荐
14+阅读 · 2018年11月29日
disentangled-representation-papers
CreateAMind
26+阅读 · 2018年9月12日
Hierarchical Imitation - Reinforcement Learning
CreateAMind
19+阅读 · 2018年5月25日
LibRec 精选:推荐的可解释性[综述]
LibRec智能推荐
10+阅读 · 2018年5月4日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
1+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2009年12月31日
Arxiv
0+阅读 · 2023年5月19日
Arxiv
31+阅读 · 2022年2月15日
A Survey on Edge Intelligence
Arxiv
52+阅读 · 2020年3月26日
Arxiv
16+阅读 · 2020年2月6日
VIP会员
最新内容
博士论文 | 面向大模型推理的内存高效算法
专知会员服务
0+阅读 · 今天15:20
美空军新型反无人机部队初探
专知会员服务
4+阅读 · 今天5:45
《防空交战流程的概率建模研究》
专知会员服务
6+阅读 · 今天5:04
ICML 2026 教程 | 数值优化理论还重要吗?
专知会员服务
4+阅读 · 7月26日
ICM 2026 | 陶哲轩:人工智能时代的数学
专知会员服务
8+阅读 · 7月26日
《反无人机交战场景下的战斗归零研究》
专知会员服务
7+阅读 · 7月26日
博士论文 | 用代码结构感知方法推进代码大模型
相关VIP内容
Meta最新WWW2022《联邦计算导论》教程,附77页ppt
专知会员服务
60+阅读 · 2022年5月5日
相关资讯
相关基金
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
1+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2009年12月31日
Top
微信扫码咨询专知VIP会员