• Home
  • Search
  • Scalable Varied Density Clustering Algorithm for Large Datasets
  • Cite Icon9
  • https://doi.org/10.4236/jsea.2010.36069Copy DOI Icon

Scalable Varied Density Clustering Algorithm for Large Datasets

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Finding clusters in data is a challenging problem especially when the clusters are being of widely varied shapes, sizes, and densities. Herein a new scalable clustering technique which addresses all these issues is proposed. In data mining, the purpose of data clustering is to identify useful patterns in the underlying dataset. Within the last several years, many clustering algorithms have been proposed in this area of research. Among all these proposed methods, density clustering methods are the most important due to their high ability to detect arbitrary shaped clusters. Moreover these methods often show good noise-handling capabilities, where clusters are defined as regions of typical densities separated by low or no density regions. In this paper, we aim at enhancing the well-known algorithm DBSCAN, to make it scalable and able to discover clusters from uneven datasets in which clusters are regions of homogenous densities. We achieved the scalability of the proposed algorithm by using the k-means algorithm to get initial partition of the dataset, applying the enhanced DBSCAN on each partition, and then using a merging process to get the actual natural number of clusters in the underlying dataset. This means the proposed algorithm consists of three stages. Experimental results using synthetic datasets show that the proposed clustering algorithm is faster and more scalable than the enhanced DBSCAN counterpart.

Loading PDF

Similar Papers
  • Research Article
  • Citations29

Data Clustering with Actuarial Applications

  • Jun 14, 2019
  • North American Actuarial Journal
  • Guojun Gan +1
  • Conference Article
  • Citations3

A Subtractive Based Subspace Clustering Algorithm on High Dimensional Data

  • Dec 01, 2009
  • Ying Deng +2
  • Conference Article
  • Citations3

Clustering data and imprecise concepts

  • Jun 01, 2011
  • Weifeng Zhang +1
  • Conference Article
  • Citations7

An Optimized Chameleon Algorithm based on Local Features

  • Feb 26, 2018
  • Xiaoxiao Cao +5
  • Conference Article

Clustering Ensemble of Clustering Algorithm for Large Datasets

  • Aug 01, 2013
  • Bencheng Yu +3
  • PDF
  • Research Article
  • Citations64

Scalable Clustering Algorithms for Big Data: A Review

  • Jan 01, 2021
  • IEEE Access
  • Mahmoud A Mahdi +2
  • Book Chapter
  • Citations10

An Optimized Approach for Density Based Spatial Clustering Application with Noise

  • Jan 01, 2014
  • Rakshit Arya +1
  • PDF
  • Research Article
  • Citations3

A Graph Convolution Network Based on Improved Density Clustering for Recommendation System

  • Mar 26, 2022
  • Information Technology and Control
  • Yue Li
  • Research Article
  • Citations5

Reclust: an efficient clustering algorithm for mixed data based on reclustering and cluster validation

  • Jan 01, 2022
  • Indonesian Journal of Electrical Engineering and Computer Science
  • Amala Jayanthi Maria Soosai Arockiam +1
  • Single Book
  • Citations10

Data Clustering

  • Aug 17, 2022
  • Tang, Niansheng
  • Conference Article
  • Citations10

Theoretical analysis of the Minimum Sum of Squared Similarities sampling for Nyström-based spectral clustering

  • Jul 01, 2016
  • Djallel Bouneffouf +1
  • Conference Article
  • Citations6

Scalable evolutionary clustering algorithm with Self Adaptive Genetic Operators

  • Jul 01, 2010
  • Elizabeth Leon +2
  • Research Article
  • Citations7

PFHC: A clustering algorithm based on data partitioning for unevenly distributed datasets

  • Nov 27, 2008
  • Fuzzy Sets and Systems
  • Yihong Dong +4
  • Research Article
  • Citations3

Simrec: a similarity measure recommendation system for mixed data clustering algorithms

  • Feb 22, 2025
  • Journal of Big Data
  • Abdoulaye Diop +5
  • Research Article

KNCM: Kernel Neutrosophic c-Means Clustering

  • Nov 01, 2017
  • Zenodo (CERN European Organization for Nuclear Research)
  • Yaman Akbulut +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.