arrow
Return

Negative results for software effort estimation

delete2016-11-21
delete40
PRE
AI
T
Tim Menzies *
Y
Ye Yang
G
George Mathew
B
Barry Boehm
J
Jairus Hihn
DOI:10.1007/s10664-016-9472-2delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
More than half the literature on software effort estimation (SEE) focuses on comparisons of new estimation methods. Surprisingly, there are no studies comparing state of the art latest methods with decades-old approaches. Accordingly, this paper takes five steps to check if new SEE methods generated better estimates than older methods. Firstly, collect effort estimation methods ranging from classical COCOMO (parametric estimation over a pre-determined set of attributes) to modern (reasoning via analogy using spectral-based clustering plus instance and feature selection, and a recent baseline method proposed in ACM Transactions on Software Engineering). Secondly, catalog the list of objections that lead to the development of post-COCOMO estimation methods. Thirdly, characterize each of those objections as a comparison between newer and older estimation methods. Fourthly, using four COCOMO-style data sets (from 1991, 2000, 2005, 2010) and run those comparisons experiments. Fifthly, compare the performance of the different estimators using a Scott-Knott procedure using (i) the A12 effect size to rule out small differences and (ii) a 99 % confident bootstrap procedure to check for statistically different groupings of treatments. The major negative result of this paper is that for the COCOMO data sets, nothing we studied did any better than Boehms original procedure. Hence, we conclude that when COCOMO-style attributes are available, we strongly recommend (i) using that data and (ii) use COCOMO to generate predictions. We say this since the experiments of this paper show that, at least for effort estimation, how data is collected is more important than what learner is applied to that data.
Keywords:
Effort estimation
COCOMO
CART
Nearest neighbor
Clustering
Feature selection
Prototype generation
Bootstrap sampling
Effect size
A12
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Empirical Software Engineering cover
Empirical Software Engineering
IF:
3.6
Papers:
2.0K
Citations:
5.3K

Organization

U
university of southern california
Scholars:
4.6W
Papers: 3.8W
Citations: 51
N
national aeronautics & space administration (nasa)
Scholars:
3.1W
Papers: 2.6W
Citations: 46
S
Stevens Institute of Technology
Scholars:
2.9K
Papers: 2.9K
Citations: 3.2K
N
North Carolina State University
Scholars:
2.6W
Papers: 2.3W
Citations: 3.7W
researcher View more organizations