English

Test cases as a measurement instrument in experimentation

Software Engineering 2022-04-26 v2

Abstract

Background: Test suites are frequently used to quantify relevant software attributes, such as quality or productivity. Problem: We have detected that the same response variable, measured using different test suites, yields different experiment results. Aims: Assess to which extent differences in test case construction influence measurement accuracy and experimental outcomes. Method: Two industry experiments have been measured using two different test suites, one generated using an ad-hoc method and another using equivalence partitioning. The accuracy of the measures has been studied using standard procedures, such as ISO 5725, Bland-Altman and Interclass Correlation Coefficients. Results: There are differences in the values of the response variables up to +-60%, depending on the test suite (ad-hoc vs. equivalence partitioning) used. Conclusions: The disclosure of datasets and analysis code is insufficient to ensure the reproducibility of SE experiments. Experimenters should disclose all experimental materials needed to perform independent measurement and re-analysis.

Keywords

Cite

@article{arxiv.2111.05287,
  title  = {Test cases as a measurement instrument in experimentation},
  author = {Oscar Dieste and Fernando Uyaguari and Sira Vegas and Natalia Juristo},
  journal= {arXiv preprint arXiv:2111.05287},
  year   = {2022}
}

Comments

Author list fixed

R2 v1 2026-06-24T07:32:40.593Z