Cross-company defect prediction via semi-supervised clustering-based data filtering and MSTrA-based transfer learning

Xiao Yu,Kwabena Ebo Bennin,Man Wu,Mandi Fu,Chuanxiang Ma,Yiheng Jian

doi:10.1007/s00500-018-3093-1

Abstract

Cross-company defect prediction (CCDP) is a practical way that trains a prediction model by exploiting one or multiple projects of a source company and then applies the model to a target company. Unfortunately, larger irrelevant cross-company (CC) data usually make it difficult to build a prediction model with high performance. On the other hand, brute force leveraging of CC data poorly related to within-company data may decrease the prediction model performance. To address such issues, we aim to provide an effective solution for CCDP. First, we propose a novel semi-supervised clustering-based data filtering method (i.e., SSDBSCAN filter) to filter out irrelevant CC data. Second, based on the filtered CC data, we for the first time introduce multi-source TrAdaBoost algorithm, an effective transfer learning method, into CCDP to import knowledge not from one but from multiple sources to avoid negative transfer. Experiments on 15 public datasets indicate that: (1) our proposed SSDBSCAN filter achieves better overall performance than compared data filtering methods; (2) our proposed CCDP approach achieves the best overall performance among all tested CCDP approaches; and (3) our proposed CCDP approach performs significantly better than with-company defect prediction models.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Cross-company defect prediction via semi-supervised clustering-based data filtering and MSTrA-based transfer learning

Abstract

Talk to us

Similar Papers

More From: Soft Computing

Lead the way for us

Journal: Soft Computing	Publication Date: Mar 8, 2018
Citations: 41

Similar Papers

A Multi-Source TrAdaBoost Approach for Cross-Company Defect Prediction
Xiao Yu ... Guoping Nie
-
Xiao Yu, et. al.Xiao Yu ... Guoping Nie
01 Jul 2016
01 Jul 2016

Improving Cross-Company Defect Prediction with Data Filtering
Xiao Yu ... Weiqiang Peng
International Journal of Software Engineering and Knowledge Engineering | VOL. 27
Xiao Yu, et. al.Xiao Yu ... Weiqiang Peng
01 Nov 2017
International Journal of Software Engineering and Knowledge Engineering | VOL. 27

Object-Oriented Metrics for Defect Prediction
Satwinder Singh ... Rozy Singla
-
Satwinder Singh, et. al.Satwinder Singh ... Rozy Singla
13 Jun 2018
13 Jun 2018

Heterogeneous cross-company defect prediction by unified metric representation and CCA-based transfer learning
Xiaoyuan Jing ... Fei Wu
-
Xiaoyuan Jing, et. al.Xiaoyuan Jing ... Fei Wu
30 Aug 2015
30 Aug 2015

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Cross-company defect prediction via semi-supervised clustering-based data filtering and MSTrA-based transfer learning

Abstract

Talk to us

Similar Papers

More From: Soft Computing