• Home
  • Search
  • InitKmix -- A Novel Initial Partition Generation Algorithm for Clustering Mixed Data using k-means-based Clustering
  • https://doi.org/10.13140/rg.2.2.21979.62244Copy DOI Icon

InitKmix -- A Novel Initial Partition Generation Algorithm for Clustering Mixed Data using k-means-based Clustering

  • Jul 22, 2020
  • Amir Ahmad +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Mixed datasets consist of both numeric and categorical attributes. Various k-means-based clustering algorithms have been developed for these datasets. Generally, these algorithms use random partition as a starting point, which tends to produce different clustering results for different runs. In this paper, we propose, initKmix, a novel algorithm for finding an initial partition for k-means-based clustering algorithms for mixed datasets. In the initKmix algorithm, a k-means-based clustering algorithm is run many times, and in each run, one of the attributes is used to create initial clusters for that run. The clustering results of various runs are combined to produce the initial partition. This initial partition is then used as a seed to a k-means-based clustering algorithm to cluster mixed data. Experiments with various categorical and mixed datasets showed that initKmix produced accurate and consistent results, and outperformed the random initial partition method and other state-of-the-art initialization methods. Experiments also showed that k-means-based clustering for mixed datasets with initKmix performed similar to or better than many state-of-the-art clustering algorithms for categorical and mixed datasets.

Similar Papers
  • Research Article
  • Citations6

Clustering algorithm for mixed datasets using density peaks and Self-Organizing Generative Adversarial Networks

  • Jun 05, 2020
  • Chemometrics and Intelligent Laboratory Systems
  • K Balaji +2
  • Research Article
  • Citations3

Simrec: a similarity measure recommendation system for mixed data clustering algorithms

  • Feb 22, 2025
  • Journal of Big Data
  • Abdoulaye Diop +5
  • Book Chapter
  • Citations8

Parameter Free Mixed-Type Density-Based Clustering

  • Jan 01, 2018
  • Sahar Behzadi +2
  • PDF
  • Research Article
  • Citations14

Clustering of mixed-type data considering concept hierarchies: problem specification and algorithm

  • Apr 25, 2020
  • International Journal of Data Science and Analytics
  • Sahar Behzadi +3
  • Research Article
  • Citations5

Reclust: an efficient clustering algorithm for mixed data based on reclustering and cluster validation

  • Jan 01, 2022
  • Indonesian Journal of Electrical Engineering and Computer Science
  • Amala Jayanthi Maria Soosai Arockiam +1
  • Book Chapter
  • Citations6

Unsupervised Classification of Mixed Data Type of Attributes Using Genetic Algorithm (Numeric, Categorical, Ordinal, Binary, Ratio-Scaled)

  • Jan 01, 2014
  • Rohit Rastogi +4
  • Research Article
  • Citations6

GRID distribution supports clustering validation of large mixed microarray data sets

  • May 12, 2011
  • EMBnet.journal
  • Angelica Tulipano +6
  • Research Article
  • Citations15

Modified FDP cluster algorithm and its application in protein conformation clustering analysis

  • May 20, 2019
  • Digital Signal Processing
  • Guiyan Wang +2
  • Conference Article
  • Citations2

An enhanced adjacent partitioning PTS technique with low computational complexity

  • Jul 01, 2007
  • Jae-Kwon Lee +2
  • Book Chapter
  • Citations45

A New Feature Weighted Fuzzy Clustering Algorithm

  • Jan 01, 2005
  • Jie Li +2
  • Research Article
  • Citations5

A heuristic algorithm for power-network clustering

  • Apr 01, 1992
  • Canadian Journal of Electrical and Computer Engineering
  • Hesham K Temraz +1
  • Conference Article
  • Citations2007

CURE

  • Jun 01, 1998
  • Sudipto Guha +2
  • Research Article
  • Citations5

A Combined Clustering Algorithm Based on ESynC Algorithm and a Merging Judgement Process of Micro-Clusters

  • May 27, 2021
  • International Journal of Uncertainty, Fuzziness and Knowledge-Based Systems
  • Xinquan Chen +1
  • Book Chapter

A Seed-Based Inter-Domain Supervised Framework to Cluster Mixed Data Types

  • Jan 01, 2013
  • Artur Abdullin +1
  • Research Article
  • Citations5

Cuckoo Search based K-Prototype Clustering Algorithm

  • Jan 01, 2017
  • Asian Journal of Research in Social Sciences and Humanities
  • K Lakshmi +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.