Published March 30, 2020 | Version v2
Conference paper Open

Investigation of Dataset Features for Just-in-Time Defect Prediction

  • 1. De La Salle University

Description

Just-in-time (JIT) defect prediction refers to the technique of predicting whether a code change is defective. Many contributions
have been made in this area through the excellent dataset by Kamei. In this paper, we revisit the dataset and highlight preprocessing
difficulties with the dataset and the limitations of the dataset on unsupervised learning. Secondly, we propose certain features in
the Kamei dataset that can be used for training models. Lastly, we discuss the limitations of the dataset’s features.

Files

csc901dpaper-formatted.pdf

Files (419.3 kB)

Name Size Download all
md5:d59412ad35385ca08a6fad1f3883fcbf
419.3 kB Preview Download