The Community for Technology Leaders
RSS Icon
Subscribe
Issue No.04 - April (1998 vol.24)
pp: 278-301
ABSTRACT
<p><b>Abstract</b>—This paper describes a procedure for analyzing unbalanced datasets that include many nominal- and ordinal-scale factors. Such datasets are often found in company datasets used for benchmarking and productivity assessment. The two major problems caused by lack of balance are that the impact of factors can be concealed and that spurious impacts can be observed. These effects are examined with the help of two small artificial datasets. The paper proposes a method of forward pass residual analysis to analyze such datasets. The analysis procedure is demonstrated on the artificial datasets and then applied to the COCOMO dataset. The paper ends with a discussion of the advantages and limitations of the analysis procedure.</p>
INDEX TERMS
Software metrics, statistical analysis, unbalanced datasets, benchmarking data, analysis of variance, residual analysis.
CITATION
Barbara Kitchenham, "A Procedure for Analyzing Unbalanced Datasets", IEEE Transactions on Software Engineering, vol.24, no. 4, pp. 278-301, April 1998, doi:10.1109/32.677185
31 ms
(Ver 2.0)

Marketing Automation Platform Marketing Automation Tool