• DocumentCode
    3187176
  • Title

    Data collection for Software Defect Prediction - An exploratory case study of open source software projects

  • Author

    Mausa, Goran ; Grbac, Tihana Galinac ; Basic, Bojana Dalbelo

  • Author_Institution
    Fac. of Eng., Univ. of Rijeka, Rijeka, Croatia
  • fYear
    2015
  • fDate
    25-29 May 2015
  • Firstpage
    463
  • Lastpage
    469
  • Abstract
    Software Defect Prediction (SDP) empirical studies are highly biased with the quality of data and widely suffer from limited generalizations. The main reasons are the lack of data and its systematic data collection procedures. Our research aims at producing the first systematically defined data collection procedure for SDP datasets that are obtained by linking separate development repositories. This paper is the first step to achieving that objective, performing an exploratory study. We review the existing literature on approaches and tools used in the collection of SDP datasets, derive a detailed collection procedure and test it in this exploratory study. We quantify the bias that may be caused by the issues we identified and we review 35 tools for software product metrics collection. The most critical issues are many-to-many relation between bug-file links, duplicated bug-file links and the issue of untraceable bugs. Our research provides more detailed, experience based data collection procedure, crucial for further development of SDP body of knowledge. Furthermore, our findings enabled us to develop the automatic data collection tool.
  • Keywords
    program testing; project management; public domain software; software metrics; SDP datasets; automatic data collection tool; data quality; duplicated bug-file links; experience based data collection procedure; many-to-many relation; open source software projects; software defect prediction; software product metrics collection; systematic data collection procedures; Computer bugs; Data collection; Joining processes; Software; Software metrics; Systematics;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Information and Communication Technology, Electronics and Microelectronics (MIPRO), 2015 38th International Convention on
  • Conference_Location
    Opatija
  • Type

    conf

  • DOI
    10.1109/MIPRO.2015.7160316
  • Filename
    7160316