Comparative evaluation of region query strategies for DBSCAN clustering

Severino F Galán

doi:10.1016/j.ins.2019.06.036

Abstract

Clustering is a technique that allows data to be organized into groups of similar objects. DBSCAN (Density-Based Spatial Clustering of Applications with Noise) constitutes a popular clustering algorithm that relies on a density-based notion of cluster and is designed to discover clusters of arbitrary shape. The computational complexity of DBSCAN is dominated by the calculation of the ϵ-neighborhood for every object in the dataset. Thus, the efficiency of DBSCAN can be improved in two different ways: (1) by reducing the overall number of ϵ-neighborhood queries (also known as region queries), or (2) by reducing the complexity of the nearest neighbor search conducted for each region query. This paper deals with the first issue by considering the most relevant region query strategies for DBSCAN, all of them characterized by inspecting the neighborhoods of only a subset of the objects in the dataset. We comparatively evaluate these region query strategies (or DBSCAN variants) in terms of clustering effectiveness and efficiency; additionally, a novel region query strategy is introduced in this work. The results show that some DBSCAN variants are only slightly inferior to DBSCAN in terms of effectiveness, while greatly improving its efficiency. Among these variants, the novel one outperforms the rest.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Comparative evaluation of region query strategies for DBSCAN clustering

Abstract

Talk to us

Similar Papers

More From: Information Sciences

Lead the way for us

Journal: Information Sciences	Publication Date: Jun 13, 2019
Citations: 46

Similar Papers

GNN-DBSCAN: A new density-based algorithm using grid and the nearest neighbor
Li Yihong ... Lan Xiaolong
Journal of Intelligent & Fuzzy Systems | VOL. 41
Li Yihong, et. al.Li Yihong ... Lan Xiaolong
16 Dec 2021
Journal of Intelligent & Fuzzy Systems | VOL. 41

Particle swarm Optimized Density-based Clustering and Classification: Supervised and unsupervised learning approaches
Chun Guan ... Kevin Kam Fung Yuen
Swarm and Evolutionary Computation | VOL. 44
Chun Guan, et. al.Chun Guan ... Kevin Kam Fung Yuen
10 Oct 2018
Swarm and Evolutionary Computation | VOL. 44

Exploiting Taxi Demand Hotspots Based on Vehicular Big Data Analytics
Lu Zhang ... Cailian Chen
-
Lu Zhang, et. al.Lu Zhang ... Cailian Chen
01 Sep 2016
01 Sep 2016

Star Catalog Generation for Satellite Attitude Navigation Using Density Based Clustering
Muhammad Arif Saifudin ... Bib Paruhum Silalahi
Journal of Computer Science | VOL. 11
Muhammad Arif Saifudin, et. al.Muhammad Arif Saifudin ... Bib Paruhum Silalahi
01 Dec 2015
Journal of Computer Science | VOL. 11

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Comparative evaluation of region query strategies for DBSCAN clustering

Abstract

Talk to us

Similar Papers

More From: Information Sciences