A case study on partitioning data for classification

Bikash Kanti Sarkar

doi:10.1504/ijids.2016.075788

A case study on partitioning data for classification

Bikash Kanti Sarkar

https://doi.org/10.1504/ijids.2016.075788

Copy DOI

Journal: International Journal of Information and Decision Sciences	Publication Date: Jan 1, 2016
Citations: 6

Affiliation: Birla Institute of Technology, Mesra

#Number Of Samples #Classification Model + Show 8 more

Abstract
Full-Text
Similar Papers

Abstract

Designing accurate model for classification problem is a real concern in context of machine learning. The various factors such as inclusion of excellent samples in the training set, the number of samples as well as the proportion of each class type in the set (that would be sufficient for designing model) play important roles in this purpose. In this article, an investigation is introduced to address the question of what proportion of the samples should be devoted to the training set for developing a better classification model. The experimental results on several datasets, using C4.5 classifier, shows that any equidistributed data partitioning in between (20%, 80%) and (30%, 70%) may be considered as the best sample partition to build classification model irrespective to domain, size and class imbalanced.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: International Journal of Information and Decision Sciences

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.