Classification Breast Cancer Revisited with Machine Learning

  • Parhusip H
  • Susanto B
  • Linawati L
  • et al.
N/ACitations
Citations of this article
27Readers
Mendeley users who have this article in their library.

Abstract

The article presents the study of several machine learning algorithms that are used to study breast cancer data with 33 features from 569 samples. The purpose of this research is to investigate the best algorithm for classification of breast cancer. The data may have different scales with different large range one to the other features and hence the data are transformed before the data are classified. The used classification methods in machine learning are logistic regression, k-nearest neighbor, Naive bayes classifier, support vector machine, decision tree and random forest algorithm. The original data and the transformed data are classified with size of data test is 0.3. The SVM and Naive Bayes algorithms have no improvement of accuracy with random forest gives the best accuracy among all. Therefore the size of data test is reduced to 0.25 leading to improve all algorithms in transformed data classifications. However, random forest algorithm still gives the best accuracy.

Cite

CITATION STYLE

APA

Parhusip, H. A., Susanto, B., Linawati, L., Trihandaru, S., Sardjono, Y., & Mugirahayu, A. S. (2020). Classification Breast Cancer Revisited with Machine Learning. International Journal of Data Science, 1(1), 42–50. https://doi.org/10.18517/ijods.1.1.42-50.2020

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free