Comparison of Machine Learning Algorithms for Spam Detection

10Citations
Citations of this article
34Readers
Mendeley users who have this article in their library.
Get full text

Abstract

The Internet is used as a tool to offer people with endless knowledge. It is a global platform which is used for connectivity, communication, and sharing. At almost no cost, an individual can use the Internet to send email messages, update tweets, and Facebook messages to a vast number of people. These messages can also contain unsolicited advertisement which is identified as a spam. The company Twitter too is massively affected by spamming and it is an alarming issue for them. Twitter considers spam as actions that are unsolicited and repeated. These include tweet repetition, and the URLs that lead users to completely unrelated websites. The authors’ have worked with twitter’s dataset focusing on tweets about “iPhone”. It was collected by using an API which was further pre-processed. In this paper, content-based features have been selected that recognize the spamming tweet by using R. Multiple machine learning algorithms were applied to detect spamming tweets: Naive Bayes, Logistic Regression, KNN, Decision Tree, and Support Vector Machine. It was observed that the best performance was achieved by Naive Bayes Algorithm giving an accuracy of 89%.

Cite

CITATION STYLE

APA

Sadia, A., Bashir, F., Khan, R. Q., Bashir, A., & Khalid, A. (2023). Comparison of Machine Learning Algorithms for Spam Detection. Journal of Advances in Information Technology, 14(2), 178–184. https://doi.org/10.12720/jait.14.2.178-184

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free