Naive Bayesian Spam Filtering

Yuhao Zhu

doi:10.54097/hset.v38i.5734

Abstract

The spam filtering system is used to identify which emails in the received emails are completely meaningless to the recipient and perform operations such as interception and deletion. Nowadays, with the rapid development of the Internet, while e-mail provides convenience for people, spam also comes along with it, which brings many troubles to users. According to statistics, 80% of the emails in the world are spam, and e-spam is really annoying. Therefore, how to solve the problem of filtering emails has important practical significance. Spam filtering using Bayesian theory is a statistical technique applied to email filtering. It essentially uses Bayesian classification to discriminate the attributes of emails, including spam and non-spam. Bayesian-based spam filtering is a very effective technique that can modify the model to meet the needs of specific users and give a lower spam detection rate that is acceptable to users. In this experiment, we use Naive Bayes for experiments, and we use Unigram and bigram methods to preprocess the data, respectively. Finally, it is concluded that the data processing accuracy of unigram and bigram is greater than 0.75, and bigram performs better in four different evaluation indicators.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Naive Bayesian Spam Filtering

Abstract

Talk to us

Similar Papers

More From: Highlights in Science, Engineering and Technology

Lead the way for us

Journal: Highlights in Science, Engineering and Technology	Publication Date: Mar 16, 2023
License type: CC BY-NC 4.0

Similar Papers

Towards the Low False Alarms and High Detection Rate in Intrusions Detection System
Hafiz Muhammad Imran ... Azween Bin Abdullah
International Journal of Machine Learning and Computing | VOL. -
Hafiz Muhammad Imran, et. al.Hafiz Muhammad Imran ... Azween Bin Abdullah
01 Jan 2013
International Journal of Machine Learning and Computing | VOL. -

Is it possible to identify factors which affect the efficacy of screening for congenital malformations by ultrasonography?
A Tabor ... B L Pedersen
Ultrasound in Obstetrics & Gynecology | VOL. 18
A Tabor, et. al.A Tabor ... B L Pedersen
01 Oct 2001
Ultrasound in Obstetrics & Gynecology | VOL. 18

A new approach to evaluating earnings management models
June Woo Park ... Giseok Nam
International Journal of Managerial and Financial Accounting | VOL. 8
June Woo Park, et. al.June Woo Park ... Giseok Nam
01 Jan 2015
International Journal of Managerial and Financial Accounting | VOL. 8

An Improved Intrusion Detection based on Neural Network and Fuzzy Algorithm
Liang He
Journal of Networks | VOL. 9
Liang HeLiang He
08 May 2014
Journal of Networks | VOL. 9

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Naive Bayesian Spam Filtering

Abstract

Talk to us

Similar Papers

More From: Highlights in Science, Engineering and Technology