Impact of preprocessing on medical data classification

Sarab Almuhaideb,Mohamed El Bachir Menai

doi:10.1007/s11704-016-5203-5

Impact of preprocessing on medical data classification

Sarab Almuhaideb, Mohamed El Bachir Menai

https://doi.org/10.1007/s11704-016-5203-5

Copy DOI

Journal: Frontiers of Computer Science	Publication Date: Oct 4, 2016
Citations: 25

Affiliation: Prince Sultan University

#Medical Data Classification #Ant Colony Optimization + Show 8 more

Abstract
Full-Text
Similar Papers

Abstract

The significance of the preprocessing stage in any data mining task is well known. Before attempting medical data classification, characteristics ofmedical datasets, including noise, incompleteness, and the existence of multiple and possibly irrelevant features, need to be addressed. In this paper, we show that selecting the right combination of preprocessing methods has a considerable impact on the classification potential of a dataset. The preprocessing operations considered include the discretization of numeric attributes, the selection of attribute subset(s), and the handling of missing values. The classification is performed by an ant colony optimization algorithm as a case study. Experimental results on 25 real-world medical datasets show that a significant relative improvement in predictive accuracy, exceeding 60% in some cases, is obtained.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: Frontiers of Computer Science

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.