Abstract
This research aims to classify Short Message Service (SMS) data by applying classification models that have studied SMS data to classify SMS data into SMS spam and SMS ham. The classification model is made from data mining algorithms: Naive Bayes and support vector machine. Before implementing the two algorithms, the SMS data will go through a text preprocessing stage, including data cleaning (whitespace removal, removal of punctuation, and removal of numbers), case folding, stemming, tokenizing, and stop word removal. In this research, a comparison of the accuracy of the two data mining methods will be carried out to see and get the best classification algorithm. Researchers also implemented several experiments by comparing the use of testing data by 20 and 30% and comparing the application of preprocessing stemming and without stemming. This study found that the support vector machine algorithm using testing data of 20% by applying the stemming stage had the highest accuracy rate, 97.5%.
Author supplied keywords
Cite
CITATION STYLE
Madyatmadja, E. D., Aldi, Fheren, F., Angelica, H., Juwitasary, H., & Sembiring, D. J. M. (2023). Comparative Study: Algorithms for Short Message Service Classification. Journal of Computer Science, 19(11), 1333–1344. https://doi.org/10.3844/jcssp.2023.1333.1344
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.