返回
Automated training-set creation for software architecture traceability problem
DOI:10.1007/s10664-016-9476-y.png)
摘要
En 中文
Automated trace retrieval methods based on machine-learning algorithms can significantly reduce the cost and effort needed to create and maintain traceability links between requirements, architecture and source code. However, there is always an upfront cost to train such algorithms to detect relevant architectural information for each quality attribute in the code. In practice, training supervised or semi-supervised algorithms requires the expert to collect several files of architectural tactics that implement a quality requirement and train a learning method. Establishing such a training set can take weeks to months to complete. Furthermore, the effectiveness of this approach is largely dependent upon the knowledge of the expert. In this paper, we present three baseline approaches for the creation of training data. These approaches are (i) Manual Expert-Based, (ii) Automated Web-Mining, which generates training sets by automatically mining tactic's APIs from technical programming websites, and lastly (iii) Automated Big-Data Analysis, which mines ultra-large scale code repositories to generate training sets. We compare the trace-link creation accuracy achieved using each of these three baseline approaches and discuss the costs and benefits associated with them. Additionally, in a separate study, we investigate the impact of training set size on the accuracy of recovering trace links. The results indicate that automated techniques can create a reliable training set for the problem of tracing architectural tactics.
Keyword:
Architecture traceability
Dataset generation
Architecturally significant requirements
Automation
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.6
论文数:
2.0K
被引数:
5.3K
机构
引用论文
Bilateral Wilms Tumor in a Boy with Severe Hypospadias and Cryptorchidism Due to a Heterozygous Mutation in the WT1 Gene由于WT1基因的杂合突变,患有严重尿道下裂和隐睾的男孩的双侧Wilms肿瘤
Impaired Spermatogenesis and gr/gr Deletions Related to Y Chromosome Haplogroups in Korean Men
PLoS ONE
IF0
Using evolutionary algorithms as instance selection for data reduction in KDD: An experimental study使用进化算法作为KDD中数据约简的实例选择: 一项实验研究

