[Author: Chan, Victor K Y] AND [Keyword: Software Metrics] : Search

Anywhere

Advanced Search

SEARCH GUIDE

Results: 1 - 1of1

Follow results:

refine search

Filters

per page:

Sort: Relevance

Context for search term 1Search term 1*

All Dates

LastSelect static range

Custom Range

Select starting monthSelect starting year

Select ending monthSelect ending year

Advanced

Search name	Searched On	Run search
[Author: Macdonald, Alexander E] AND [Keyword: Generation] (1)	14 Mar 2025	Run
[Author: Wilson, David] AND [Keyword: Recommender Systems] (2)	14 Mar 2025	Run
[Author: Chan, Victor K Y] AND [Keyword: Software Metrics] (1)	14 Mar 2025	Run
[Author: Tonelli, Roberto] AND [Keyword: Software Metrics] (2)	14 Mar 2025	Run
[Author: Hudepohl, John P] AND [Keyword: Software Quality] (2)	14 Mar 2025	Run

articleNo Access
A STATISTICAL METHODOLOGY TO SIMPLIFY SOFTWARE METRIC MODELS CONSTRUCTED USING INCOMPLETE DATA SAMPLES
- VICTOR K. Y. CHAN,
- W. ERIC WONG, and
- T. F. XIE
International Journal of Software Engineering and Knowledge Engineering01 Dec 2007
Preview Abstract
Software metric models predict the target software metric(s), e.g., the development work effort or defect rates, for any future software project based on the project's predictor software metric(s), e.g., the project team size. Obviously, the construction of such a software metric model makes use of a data sample of such metrics from analogous past projects. However, incomplete data often appear in such data samples. Moreover, the decision on whether a particular predictor metric should be included is most likely based on an intuitive or experience-based assumption that the predictor metric has an impact on the target metric with a statistical significance. However, this assumption is usually not verifiable "retrospectively" after the model is constructed, leading to redundant predictor metric(s) and/or unnecessary predictor metric complexity. To solve all these problems, we derived a methodology consisting of the k-nearest neighbors (k-NN) imputation method, statistical hypothesis testing, and a "goodness-of-fit" criterion. This methodology was tested on software effort metric models and software quality metric models, the latter usually suffers from far more serious incomplete data. This paper documents this methodology and the tests on these two types of software metric models.