Software quality assurance remains a major challenge in industrial environments, where large-scale and long-lived systems inevitably accumulate defects. Identifying the location of a fault is often time-consuming and costly, particularly during maintenance phases when developers must rely primarily on textual bug reports rather than complete runtime or code-level context. In this study, we investigated if artificial intelligence can support fault localization using only the natural-language content of bug reports. By relying only on textual information, our approach requires no access to source code, execution traces, or static analysis artifacts, making it directly deployable within existing industrial maintenance workflows. We framed fault localization as a supervised text classification problem and evaluated three traditional machine learning models (Logistic Regression, Support Vector Machine, and Random Forest) and two fine-tuned transformer-based language models (RoBERTa-Base and Distil-RoBERTa). Our evaluation used proprietary data from ABB Robotics in Sweden, comprising five years of resolved industrial bug reports, each linked to its verified code fix. This setting allowed us to assess model effectiveness under realistic industrial constraints. Our results showed that traditional models using term frequency-inverse document features consistently outperformed the fine-tuned language models on this dataset, while data augmentation improved Random Forest performance. These findings challenge the assumption that transformer-based models universally outperform classical approaches in industrial contexts with domain-specific data. We demonstrated that historical bug reports can be systematically used for text-based, artificial intelligence-assisted fault localization, providing a scalable, low-cost, and empirically grounded complement to common debugging practices in industry.


翻译:软件质量保障在工业环境中仍然是一个重大挑战,大规模且长期运行的系统不可避免地会积累缺陷。定位故障的位置通常耗时且成本高昂,尤其在维护阶段,开发者主要依赖文本形式的Bug报告,而非完整的运行时或代码级上下文。在本研究中,我们探究了人工智能是否能够仅利用Bug报告中的自然语言内容来支持故障定位。由于仅依赖文本信息,我们的方法无需访问源代码、执行轨迹或静态分析产物,因此可直接部署于现有的工业维护工作流程。我们将故障定位定义为有监督文本分类问题,并评估了三种传统机器学习模型(逻辑回归、支持向量机和随机森林)以及两种基于微调Transformer的语言模型(RoBERTa-Base和Distil-RoBERTa)。我们的评估使用了来自瑞典ABB机器人的专有数据,包含五年内已解决的工业Bug报告,每条报告均关联到已验证的代码修复。这一设置使我们能够在真实工业约束下评估模型有效性。结果表明,使用词频-逆文档频率特征的传统模型在该数据集上始终优于微调后的语言模型,而数据增强提升了随机森林的性能。这些发现挑战了Transformer基模型在领域特定数据的工业环境中普遍优于传统方法的假设。我们证明了历史Bug报告可被系统地用于基于文本的、人工智能辅助的故障定位,为工业中的常规调试实践提供了一种可扩展、低成本且经验驱动的补充方案。

0
下载
关闭预览

相关内容

程序猿的天敌 有时是一个不能碰的magic
AI生成代码缺陷综述
专知会员服务
17+阅读 · 2025年12月8日
《基于大型语言模型的软件工程自动化研究》最新264页
专知会员服务
39+阅读 · 2025年7月14日
《利用人工智能预测飞机发动机故障》
专知会员服务
36+阅读 · 2024年7月8日
大型语言模型时代AIOps在故障管理中的综述
专知会员服务
44+阅读 · 2024年6月23日
人工智能企业技术岗位设置研究报告
专知会员服务
45+阅读 · 2022年2月26日
专知会员服务
14+阅读 · 2021年9月21日
专知会员服务
10+阅读 · 2021年1月31日
专知会员服务
31+阅读 · 2020年12月21日
【PHM算法】PHM算法 | 故障诊断建模方法
产业智能官
68+阅读 · 2020年3月16日
【质量检测】机器视觉表面缺陷检测综述
产业智能官
30+阅读 · 2018年9月24日
【机器视觉】表面缺陷检测:机器视觉检测技术
产业智能官
25+阅读 · 2018年5月30日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
Arxiv
0+阅读 · 6月1日
Arxiv
0+阅读 · 3月24日
VIP会员
最新内容
深入Project Maven:为何人工智能在战场上依然失灵
锻造未来士兵:外骨骼、基因工程与赛博格
专知会员服务
7+阅读 · 7月19日
《无人机蜂群通信技术研究》50页
专知会员服务
7+阅读 · 7月19日
战力倍增器:自主武器系统与乌克兰及加沙冲突
人工智能赋能战场情报:提速决策进程
专知会员服务
6+阅读 · 7月17日
《拥抱新兴技术:面向未来军官的教育革新》
专知会员服务
8+阅读 · 7月17日
相关VIP内容
AI生成代码缺陷综述
专知会员服务
17+阅读 · 2025年12月8日
《基于大型语言模型的软件工程自动化研究》最新264页
专知会员服务
39+阅读 · 2025年7月14日
《利用人工智能预测飞机发动机故障》
专知会员服务
36+阅读 · 2024年7月8日
大型语言模型时代AIOps在故障管理中的综述
专知会员服务
44+阅读 · 2024年6月23日
人工智能企业技术岗位设置研究报告
专知会员服务
45+阅读 · 2022年2月26日
专知会员服务
14+阅读 · 2021年9月21日
专知会员服务
10+阅读 · 2021年1月31日
专知会员服务
31+阅读 · 2020年12月21日
相关基金
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员