Dimensionality and data reduction in telecom churn prediction

Wei-Chao Lin,Shih-Wen Ke,Chih-Fong Tsai

doi:10.1108/k-03-2013-0045

Abstract

Purpose – Churn prediction is a very important task for successful customer relationship management. In general, churn prediction can be achieved by many data mining techniques. However, during data mining, dimensionality reduction (or feature selection) and data reduction are the two important data preprocessing steps. In particular, the aims of feature selection and data reduction are to filter out irrelevant features and noisy data samples, respectively. The purpose of this paper, performing these data preprocessing tasks, is to make the mining algorithm produce good quality mining results. Design/methodology/approach – Based on a real telecom customer churn data set, seven different preprocessed data sets based on performing feature selection and data reduction by different priorities are used to train the artificial neural network as the churn prediction model. Findings – The results show that performing data reduction first by self-organizing maps and feature selection second by principal component analysis can allow the prediction model to provide the highest prediction accuracy. In addition, this priority allows the prediction model for more efficient learning since 66 and 62 percent of the original features and data samples are reduced, respectively. Originality/value – The contribution of this paper is to understand the better procedure of performing the two important data preprocessing steps for telecom churn prediction.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Dimensionality and data reduction in telecom churn prediction

Abstract

Talk to us

Similar Papers

More From: Kybernetes

Lead the way for us

Journal: Kybernetes	Publication Date: Apr 29, 2014
Citations: 14

Similar Papers

Telecom churn prediction and used techniques, datasets and performance measures: a review
Hemlata Jain ... Sumit Srivastava
Telecommunication Systems | VOL. 76
Hemlata Jain, et. al.Hemlata Jain ... Sumit Srivastava
20 Oct 2020
Telecommunication Systems | VOL. 76

Data pre-processing by genetic algorithms for bankruptcy prediction
Chih-Fong Tsai ... Jui-Sheng Chou
-
Chih-Fong Tsai, et. al.Chih-Fong Tsai ... Jui-Sheng Chou
01 Dec 2011
01 Dec 2011

Comparison of Principal Component Analysis and Recursive Feature Elimination with Cross-Validation Feature Selection Algorithms for Customer Churn Prediction
Muhammad Afif Afdholul Matin ... Rima Tamara Aldisa
-
Muhammad Afif Afdholul Matin, et. al.Muhammad Afif Afdholul Matin ... Rima Tamara Aldisa
01 Jan 2023
01 Jan 2023

Telecom Churn Prediction Using Voting Classifier Ensemble Method and Supervised Machine Learning Techniques
O Pandithurai ... N Yuvaraj
ITM Web of Conferences | VOL. 56
O Pandithurai, et. al.O Pandithurai ... N Yuvaraj
01 Jan 2023
ITM Web of Conferences | VOL. 56

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Dimensionality and data reduction in telecom churn prediction

Abstract

Talk to us

Similar Papers

More From: Kybernetes