A Study of Machine Learning Algorithms on Email Spam Classification

N Sutta,X Zhang,Z Liu

doi:10.29007/qshd

A Study of Machine Learning Algorithms on Email Spam Classification

N Sutta, X Zhang + Show 1 more

Open Access

https://doi.org/10.29007/qshd

Copy DOI

Publication Date: Mar 9, 2020

Citations: 3

Affiliation: Southeast Missouri State University

#Separate Dataset #Email Spam Classification + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

Despite the fact that different techniques have been developed to filter spam, due to the spammer’s rapid adoption of new spam detection techniques, we are still overwhelmed with spam emails. Currently, machine learning techniques are the most effective ways to classify and filter spam emails. In this paper, a comprehensive comparison and analysis of the performance of various classification models on the 2007 TREC Public Spam Corpus are exhibited in various cases of without or with N- Grams as well as using separate or combined datasets. It is shown that the inclusion of the N-Grams in the pre-processing phase provides high accuracy results for classification models in most of the cases, and the models using the split approach with combined datasets give better results than models using the separate dataset.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.