A framework of mutual information kullback-leibler divergence based for clustering categorical data

4Citations
Citations of this article
9Readers
Mendeley users who have this article in their library.

Abstract

— Clustering is a process of grouping a set of objects into multiple clusters, so that the collection of similar objects will be grouped into the same cluster and dissimilar objects will be grouped into other clusters. Fuzzy k-means Algorithm is one of clustering algorithm by partitioning data into k clusters employing Euclidean distance as a distance function. This research discusses clustering categorical data using Fuzzy k-Means Kullback-Leibler Divergence. In the determination of the distance between data and center of cluster uses mutual information known as Kullback-Leibler Divergence distance between the joint distribution and the product distribution from two marginal distributions. Extensive theoretical analysis was performed to show the effectiveness of the proposed method. Moreover, the proposed method's comparison results with Fuzzy Centroid and Fuzzy k-Partition approaches in terms of response time and clustering accuracy were also performed employing several datasets from UCI Machine Learning. The experiment results show that the proposed Algorithm provides good results both from clustering quality and accuracy for clustering categorical data as compared to Fuzzy Centroid and Fuzzy k-Partition.

Cite

CITATION STYLE

APA

Yanto, I. T. R., Setiyowati, R., Azizah, N., & Rasyidah. (2021). A framework of mutual information kullback-leibler divergence based for clustering categorical data. International Journal on Informatics Visualization, 5(1), 11–15. https://doi.org/10.30630/joiv.5.1.462

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free