Performance evaluation of similarity measures for K-means clustering algorithm

D Usman,S.F Sani

doi:10.4314/bajopas.v12i2.21

Performance evaluation of similarity measures for K-means clustering algorithm

D Usman, S.F Sani

Open Access

https://doi.org/10.4314/bajopas.v12i2.21

Copy DOI

Journal: Bayero Journal of Pure and Applied Sciences	Publication Date: Feb 12, 2021
Citations: 6

#Clustering Criterion Function #High-dimensional Domain + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

Clustering is a useful technique that organizes a large quantity of unordered datasets into a small number of meaningful and coherent clusters. Every clustering method is based on the index of similarity or dissimilarity between data points. However, the true intrinsic structure of the data could be correctly described by the similarity formula defined and embedded in the clustering criterion function. This paper uses squared Euclidean distance and Manhattan distance to investigates the best method for measuring similarity between data objects in sparse and high-dimensional domain which is fast, capable of providing high quality clustering result and consistent. The performances of these two methods were reported with simulated high dimensional datasets.

Full Text