Publicación:
Comparative Evaluation of Region Query Strategies for DBSCAN Clustering

dc.contributor.authorFernández Galán, Severino
dc.date.accessioned2024-05-20T11:43:22Z
dc.date.available2024-05-20T11:43:22Z
dc.date.issued2019-10
dc.description.abstractClustering is a technique that allows data to be organized into groups of similar objects. DBSCAN (Density-Based Spatial Clustering of Applications with Noise) constitutes a popular clustering algorithm that relies on a density-based notion of cluster and is designed to discover clusters of arbitrary shape. The computational complexity of DBSCAN is dominated by the calculation of the ϵ-neighborhood for every object in the dataset. Thus, the efficiency of DBSCAN can be improved in two different ways: (1) by reducing the overall number of ϵ-neighborhood queries (also known as region queries), or (2) by reducing the complexity of the nearest neighbor search conducted for each region query. This paper deals with the first issue by considering the most relevant region query strategies for DBSCAN, all of them characterized by inspecting the neighborhoods of only a subset of the objects in the dataset. We comparatively evaluate these region query strategies (or DBSCAN variants) in terms of clustering effectiveness and efficiency; additionally, a novel region query strategy is introduced in this work. The results show that some specific DBSCAN variants are only slightly inferior to DBSCAN in terms of effectiveness, while greatly improving its efficiency. Among these variants, the novel one outperforms the rest.en
dc.description.versionversión final
dc.identifier.doihttps://doi.org/10.1016/j.ins.2019.06.036
dc.identifier.issn0020-0255
dc.identifier.urihttps://hdl.handle.net/20.500.14468/12467
dc.journal.titleInformation Sciences
dc.journal.volume502
dc.language.isoen
dc.publisherElsevier
dc.relation.centerE.T.S. de Ingeniería Informática
dc.relation.departmentInteligencia Artificial
dc.rightsinfo:eu-repo/semantics/openAccess
dc.rights.urihttp://creativecommons.org/licenses/by-nc-nd/4.0
dc.subject.keywordsClustering
dc.subject.keywordsDBSCAN algorithm
dc.subject.keywordsregion query strategy
dc.subject.keywordscomparative evaluation
dc.titleComparative Evaluation of Region Query Strategies for DBSCAN Clusteringes
dc.typejournal articleen
dc.typeartículoes
dspace.entity.typePublication
relation.isAuthorOfPublicationa91aef9f-537b-41ae-a323-5a166ee934f6
relation.isAuthorOfPublication.latestForDiscoverya91aef9f-537b-41ae-a323-5a166ee934f6
Archivos
Bloque original
Mostrando 1 - 1 de 1
Cargando...
Miniatura
Nombre:
Fernandez-Galan_Severino_DBSCAN-clustering.pdf
Tamaño:
5.66 MB
Formato:
Adobe Portable Document Format