arrow
Return

Evaluating subset selection methods for use case points estimation

delete2018-05-01
delete42
delete
OA
AI
R
Radek Šilhavý *
P
Petr Šilhavý
Z
Zdenka Prokopová
DOI:10.1016/j.infsof.2017.12.009delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
When the Use Case Points method is used for software effort estimation, users are faced with low model accuracy which impacts on its practical application. This study investigates the significance of using subset selection methods for the prediction accuracy of Multiple Linear Regression models, obtained by the stepwise approach. K-means, Spectral Clustering, the Gaussian Mixture Model and Moving Window are evaluated as appropriate subset selection techniques. The methods were evaluated according to several evaluation criteria and then statistically tested. Evaluation was performing on two independent datasets-which differ in project types and size. Both were cut by the hold-out method. If clustering were used, the training sets were clustered into 3 classes; and, for each of class, an independent regression model was created. These were later used for the prediction of testing sets. If Moving Window was used, then window of sizes 5, 10 and 15 were tested. The results show that clustering techniques decrease prediction errors significantly when compared to Use Case Points or moving windows methods. Spectral Clustering was selected as the best-performing solution, because it achieves a Sum of Squared Errors reduction of 32% for the first dataset, and 98% for the second dataset. The Mean Absolute Percentage Error is less than 1% for the second dataset for Spectral Clustering; 9% for moving window; and 27% for Use Case Points. When the first dataset is used, then prediction errors are significantly higher -53% for Spectral Clustering, but Use Case Points produces a 165% result. It can be concluded that this study proves subset selection techniques as a significant method for improving the prediction ability of linear regression models - which are used for software development effort prediction. It can also be concluded that the clustering method performs better than the moving window method.
Keywords:
Software Development Effort Estimation
Software size estimation
Clustering techniques
Spectral Clustering
K-means
Moving Window
Use Case Points
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Information and Software Technology cover
Information and Software Technology
IF:
4.3
Papers:
3.8K
Citations:
7.7K

Organization

T
Tomas Bata University Zlin
Scholars:
1.6K
Papers: 1.4K
Citations: 16
Cited Papers

Cited Papers

Soil metabolomics: Deciphering underground metabolic webs in terrestrial ecosystems
err2024-06-01
err0
errOAAI
errYang Song; Shi Yao; Xiaona Li; Tao Wang; Xin Jiang; Nanthi Bolan; Charles R. Warren; Trent R. Northen; Scott X. Chang
errShare
errSave
Improving the reliability of transaction identification in use cases
err2011-08-01
err22
PREAI
errOchodek, M.; Alchimowicz, B.; Jurkiewicz, J.; Nawrocki, J.
errShare
errSave
ROBUST SUBSPACE CLUSTERING
err2014-04-01
err262
errOAAI
errSoltanolkotabi, Mahdi; Elhamifar, Ehsan; Candes, Emmanuel J.
errShare
errSave
A flexible method to estimate the software development effort based on the classification of projects and localization of comparisons
err2013-02-01
err48
PREAI
errBardsiri, Vahid Khatibi; Jawawi, Dayang Norhayati Abang; Hashim, Siti Zaiton Mohd; Khatibi, Elham
errShare
errSave
Simplifying effort estimation based on Use Case Points
err2011-03-01
err91
PREAI
errOchodek, M.; Nawrocki, J.; Kwarciak, K.
errShare
errSave
Analysis of plating grain size effect on whisker
err2010-01-21
err0
PREAI
errSeung-Jung Shin; Jae-Jung Kim; Young Kap Son
errShare
errSave
researcher View more