Software defect prediction using a cost sensitive decision forest and voting, and a potential solution to the class imbalance problem

Michael J Siers,Md Zahidul Islam

doi:10.1016/j.is.2015.02.006

Abstract

Software development projects inevitably accumulate defects throughout the development process. Due to the high cost that defects can incur, careful consideration is crucial when predicting which sections of code are likely to contain defects. Classification algorithms used in machine learning can be used to create classifiers which can be used to predict defects. While traditional classification algorithms optimize for accuracy, cost-sensitive classification methods attempt to make predictions which incur the lowest classification cost. In this paper we propose a cost-sensitive classification technique called CSForest which is an ensemble of decision trees. We also propose a cost-sensitive voting technique called CSVoting in order to take advantage of the set of decision trees in minimizing the classification cost. We then investigate a potential solution to class imbalance within our decision forest algorithm. We empirically evaluate the proposed techniques comparing them with six (6) classifier algorithms on six (6) publicly available clean datasets that are commonly used in the research on software defect prediction. Our initial experimental results indicate a clear superiority of the proposed techniques over the existing ones.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Software defect prediction using a cost sensitive decision forest and voting, and a potential solution to the class imbalance problem

Abstract

Talk to us

Similar Papers

More From: Information Systems

Lead the way for us

Journal: Information Systems	Publication Date: Mar 10, 2015
Citations: 126

Similar Papers

Cost Sensitive Decision Forest and Voting for Software Defect Prediction
Michael J Siers ... Md Zahidul Islam
-
Michael J Siers, et. al.Michael J Siers ... Md Zahidul Islam
01 Jan 2014
01 Jan 2014

ForEx++: A New Framework for Knowledge Discovery from Decision Forests
Md Nasim Adnan ... Md Zahidul Islam
Australasian Journal of Information Systems | VOL. 21
Md Nasim Adnan, et. al.Md Nasim Adnan ... Md Zahidul Islam
08 Nov 2017
Australasian Journal of Information Systems | VOL. 21

Novel algorithms for cost-sensitive classification and knowledge discovery in class imbalanced datasets with an application to NASA software defects
Michael J Siers ... Md Zahidul Islam
Information Sciences | VOL. 459
Michael J Siers, et. al.Michael J Siers ... Md Zahidul Islam
14 May 2018
Information Sciences | VOL. 459

Automatic Feature Exploration and an Application in Defect Prediction
Yu Qiu ... Jing Xu
IEEE Access | VOL. 7
Yu Qiu, et. al.Yu Qiu ... Jing Xu
01 Jan 2019
IEEE Access | VOL. 7

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Software defect prediction using a cost sensitive decision forest and voting, and a potential solution to the class imbalance problem

Abstract

Talk to us

Similar Papers

More From: Information Systems