arrow
返回

How repeated data points affect bug prediction performance: A case study

delete2016-12-01
delete4
PRE
AI
M
Muhammed Maruf Öztürk *
A
Ahmet Zengi̇n
DOI:10.1016/j.asoc.2016.08.002delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In defect prediction studies, open-source and real-world defect data sets are frequently used. The quality of these data sets is one of the main factors affecting the validity of defect prediction methods. One of the issues is repeated data points in defect prediction data sets. The main goal of the paper is to explore how low-level metrics are derived. This paper also presents a cleansing algorithm that removes repeated data points from defect data sets. The method was applied on 20 data sets, including five open source sets, and area under the curve (AUC) and precision performance parameters have been improved by 4.05% and 6.7%, respectively. In addition, this work discusses how static code metrics should be used in bug prediction. The study provides tips to obtain better defect prediction results. (C) 2016 Elsevier B.V. All rights reserved.
Keyword:
Bug prediction
Repeated data
Software metricsa
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Applied Soft Computing 封面图
Applied Soft Computing
IF:
6.6
论文数:
1.4W
被引数:
4.8W

机构

暂无机构信息
引用论文

引用论文

err分享
err收藏
err分享
err收藏
err分享
err收藏
学者 查看更多内容