返回
Deep learning with class-level abstract syntax tree and code histories for detecting code modification requirements
DOI:10.1016/j.jss.2023.111851.png)
摘要
En 中文
Improving code quality is one of the most significant issues in the software industry. Deep learning is an emerging area of research for detecting code smells and addressing refactoring requirements. The aim of this study is to develop a deep learning-based system for code modification analysis to predict the locations and types of code modifications, while significantly reducing the need for manual labeling. We created an experimental dataset by collecting historical code data from open -source project repositories on the Internet. We introduce a novel class-level abstract syntax tree-based code embedding method for code analysis. A recurrent neural network was employed to effectively identify code modification requirements. Our system achieves an average accuracy of approximately 83% across different repositories and 86% for the entire dataset. These findings indicate that our system provides higher performance than the method-based and text-based code embedding approaches. In addition, we performed a comparative analysis with a static code analysis tool to justify the readiness of the proposed model for deployment. The correlation coefficient between the outputs demonstrates a significant correlation of 67%. Consequently, this research highlights that the deep learning-based analysis of code histories empowers software teams in identifying potential code modification requirements.(c) 2023 Elsevier Inc. All rights reserved.
Keyword:
Refactoring
Code smell
Recurrent neural network
Abstract syntax tree
Code embedding
期刊
IF:
4.1
论文数:
5.4K
被引数:
8.4K
机构
引用论文
The Role of Molecular and Inflammatory Indicators in the Assessment of Cognitive Dysfunction in a Mouse Model of Diabetes分子和炎症指标在糖尿病小鼠认知功能障碍评估中的作用
Comparing and experimenting machine learning techniques for code smell detection比较和实验用于代码气味检测的机器学习技术

