• Home
  • Search
  • Unsupervised Classification of Mixed Data Type of Attributes Using Genetic Algorithm (Numeric, Categorical, Ordinal, Binary, Ratio-Scaled)
  • Cite Icon6
  • https://doi.org/10.1007/978-81-322-1771-8_11Copy DOI Icon

Unsupervised Classification of Mixed Data Type of Attributes Using Genetic Algorithm (Numeric, Categorical, Ordinal, Binary, Ratio-Scaled)

  • Jan 1, 2014
  • Rohit Rastogi +4 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Data mining discloses hidden, previously unknown, and potentially useful information from large amounts of data. As comparison to the traditional statistical and machine learning data analysis techniques, data mining emphasizes to provide a convenient and complete environment for the data analysis. Data mining has become a popular technology in analyzing complex data. Clustering is one of the data mining core techniques. In the field of data mining and data clustering, it is a highly desirable task to perform cluster analysis on large data sets with mixed numeric, categorical, ordinal, and ratio-scaled with binary and nominal values. However, most already available data merging and grouping through clustering algorithms are effective for the numeric data rather than the mixed data set. For this purpose, this paper makes efforts to present a new amalgamation algorithm for these mixed data sets by modifying the common cost function, trace of the within cluster dispersion matrix. The genetic algorithm (GA) is used to optimize the new cost function to obtain valid clustering result. We can compare and analyze that the GA-based clustering algorithm is feasible for the high-dimensional data sets with mixed data values that are obtained in real life results. Core Idea of Our Paper: By this paper, we try to describe a technique for estimating the cost function metrics from mixed numeric, categorical and other type databases by using an uncertain grade-of-membership clustering model with the efficiency of Genetic Algorithm. This technique can be applied to the problem of opportunity analysis for business decision-making. This general approach could be adapted to many other applications where a decision agent needs to assess the value of items from a set of opportunities with respect to a reference set representing its business. For processing numeric attributes, instead of generalizing them, a prototype may be developed for experiments with synthetic and real data sets, and comparison with those of the traditional approaches. The results confirmed the feasibility of the framework and the superiority of the extended techniques.

Similar Papers
  • Book Chapter
  • Citations4

Optimization of the Numeric and Categorical Attribute Weights in KAMILA Mixed Data Clustering Algorithm

  • Jan 01, 2019
  • Nádia Junqueira Martarelli +1
  • PDF
  • Research Article
  • Citations80

Chemist versus Machine: Traditional Knowledge versus Machine Learning Techniques

  • Nov 09, 2020
  • Trends in Chemistry
  • Janine George +1
  • Conference Article
  • Citations2

In-situ optimization of cost function for genetic algorithm using neural networks applied to antenna design

  • Jan 01, 2001
  • Y.H Lee
  • Research Article
  • Citations2

Effects of different virtual monoenergetic CT image data on chest wall post-processing "unfolded ribs" and proposal of an algorithm improvement.

  • Aug 13, 2022
  • International Journal of Computer Assisted Radiology and Surgery
  • Florian Hagen +8
  • Research Article
  • Citations3

Simrec: a similarity measure recommendation system for mixed data clustering algorithms

  • Feb 22, 2025
  • Journal of Big Data
  • Abdoulaye Diop +5
  • Conference Article
  • Citations5

A New Range Noise Perturbation Method based on Privacy Preserving Data Mining

  • Mar 01, 2020
  • Jinzhao Shan +2
  • Research Article
  • Citations30

Optimizing flood predictions by integrating LSTM and physical-based models with mixed historical and simulated data

  • Jun 26, 2024
  • Heliyon
  • Jun Li +3
  • Front Matter
  • Citations10

Taking a byte out of big data

  • Oct 26, 2015
  • The Journal of the American Dental Association
  • Michael Glick
  • Research Article
  • Citations1

A Technical Approach on Large Data Distributed Over a Network

  • Jun 01, 2011
  • International Journal of Science and Engineering
  • G Suhasini +2
  • Research Article
  • Citations6

A Data Mining Approach for Cardiovascular Diagnosis

  • Dec 20, 2017
  • Open Computer Science
  • Joana Pereira +3
  • Conference Article
  • Citations20

Data Mining for Malicious Code Detection and Security Applications

  • Jan 01, 2009
  • Bhavani Thuraisingham
  • Conference Article
  • Citations12

Data Mining for Malicious Code Detection and Security Applications

  • Jan 01, 2009
  • Bhavani Thuraisingham
  • Research Article
  • Citations11

Assessing the Efficacy of Synthetic Optic Disc Images for Detecting Glaucomatous Optic Neuropathy Using Deep Learning.

  • Jun 03, 2024
  • Translational vision science & technology
  • Abadh K Chaurasia +4
  • Conference Article
  • Citations1

Metric estimation via a fuzzy grade-of-membership model applied to analysis of business opportunities

  • Nov 04, 2002
  • B.G Talbot +2
  • Research Article

OPTIMASI NEURAL NETWORK MENGGUNAKAN GENETIC ALGORITHM UNTUK PREDIKSI PENYAKIT DIABETES

  • Jan 01, 2015
  • Hilda Amalia
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.