• Home
  • Search
  • Global-Local Self-Attention-Based Long Short-Term Memory with Optimization Algorithm for Speaker Identification
  • Cite Icon1
  • https://doi.org/10.31436/iiumej.v26i1.3386Copy DOI Icon

Global-Local Self-Attention-Based Long Short-Term Memory with Optimization Algorithm for Speaker Identification

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Speaker identification (SI) involves recognizing a speaker from a group of unknown speakers, while speaker verification (SV) determines if a given voice sample belongs to a particular person. The main drawbacks of SI are session variability, noise in the background, and insufficient information. To mitigate the limitations mentioned above, this research proposes Global Local Self-Attention (GLSA) based Long Short-Term Memory (LSTM) with Exponential Neighborhood – Grey Wolf Optimization (EN-GWO) method for effective speaker identification using TIMIT and VoxCeleb 1 datasets. The GLSA is incorporated in LSTM, which focuses on the required data, and the hyperparameters are tuned using the EN-GWO, which enhances speaker identification performance. The GLSA-LSTM with EN-GWO method acquires an accuracy of 99.36% on the TIMIT dataset, and an accuracy of 93.45% on the VoxCeleb 1 datasets, while compared to SincNet and Generative Adversarial Network (SincGAN) and Hybrid Neural Network – Support Vector Machine (NN-SVM). ABSTRAK: Pengenalpastian pembicara (Speaker Identification, SI) melibatkan pengenalan pembicara daripada kumpulan pembicara yang tidak dikenali, manakala pengesahan pembicara (Speaker Verification, SV) menentukan sama ada sampel suara tertentu milik seseorang individu. Kekurangan utama dalam SI ialah variasi sesi, bunyi latar belakang, dan maklumat yang tidak mencukupi. Untuk mengatasi kekangan tersebut, kajian ini mencadangkan kaedah Global Local Self-Attention (GLSA) berasaskan Long Short-Term Memory (LSTM) dengan Pengoptimuman Grey Wolf Jiranan Eksponen (EN-GWO) bagi pengenalpastian pembicara yang berkesan menggunakan set data TIMIT dan VoxCeleb 1. GLSA digabungkan dalam LSTM yang memberi tumpuan pada data yang diperlukan, manakala parameter hiper ditala menggunakan EN-GWO untuk meningkatkan prestasi pengenalpastian pembicara. Kaedah GLSA-LSTM dengan EN-GWO mencapai ketepatan 99.36% pada dataset TIMIT dan ketepatan 93.45% pada dataset VoxCeleb 1, berbanding dengan SincNet dan Generative Adversarial Network (SincGAN) serta Hybrid Neural Network – Support Vector Machine (NN-SVM).

Similar Papers
  • Research Article

Video compression by hybrid neural networks (CNN+RNN(LSTM)) algorithm and generative adversarial network (GAN) algorithm

  • Sep 01, 2025
  • Al-Noor Journal of Engineering Management and Computer Science
  • Sama Jassim Mohammed +1
  • Research Article
  • Citations3

A multi-modal Parkinson’s disease diagnosis system from EEG signals and online handwritten tasks using grey wolf optimization based deep learning model

  • Sep 27, 2024
  • Biomedical Signal Processing and Control
  • Kaushal Kumar +1
  • PDF
  • Research Article
  • Citations38

Automatic grading for Arabic short answer questions using optimized deep learning model.

  • Aug 02, 2022
  • PLOS ONE
  • Mustafa Abdul Salam +2
  • Dissertation
  • Citations1

MACHINE LEARNING IN CROP CLASSIFICATION OF TEMPORAL MULTISPECTRAL SATELLITE IMAGE

  • May 24, 2019
  • Ravali Koppaka
  • Research Article
  • Citations4

Optimised autoencoder-based ensemble deep learning approaches for cyber-physical event classification utilizing synchrophasor PMU data

  • Sep 01, 2025
  • Results in Engineering
  • Dhinu Lal M +1
  • Research Article
  • Citations1

Combined use of long short‐term memory neural network and quantum computation for hierarchical forecasting of locational marginal prices

  • Feb 01, 2025
  • Energy Conversion and Economics
  • Xin Huang +6
  • PDF
  • Research Article
  • Citations4

Empirical Comparison between Deep and Classical Classifiers for Speaker Verification in Emotional Talking Environments

  • Sep 27, 2022
  • Information
  • Ali Bou Nassif +4
  • Research Article
  • Citations6

Fault Detection of Wheelset Bearings through Vibration-Sound Fusion Data Based on Grey Wolf Optimizer and Support Vector Machine

  • Aug 28, 2024
  • Technologies
  • Tianhao Wang +3
  • Research Article

<b>PERBANDINGAN ALGORITMA <i>SUPPORT VECTOR MACHINE </i>DAN <i>LONG SHORT-TERM MEMORY </i>UNTUK KLASIFIKASI EMOSI MAHASISWA PADA PLATFORM <i>X </i> </b>

  • Apr 07, 2026
  • Jurnal Inovasi Pendidikan dan Teknologi Informasi (JIPTI)
  • Nazilatul Azza +1
  • Conference Article
  • Citations15

Compensating for Mismatch in High-Level Speaker Recognition

  • Jun 01, 2006
  • W Campbell
  • PDF
  • Research Article
  • Citations10

The Research of Air Combat Intention Identification Method Based on BiLSTM + Attention

  • Jun 12, 2023
  • Electronics
  • Bin Tan +3
  • Research Article
  • Citations8

Categorizing Natural Language-Based Customer Satisfaction: An Implementation Method Using Support Vector Machine and Long Short-Term Memory Neural Network

  • May 02, 2021
  • International Journal of Integrated Engineering
  • Ralph Sherwin A Corpuz
  • PDF
  • Research Article
  • Citations30

Analysis and Application of Grey Wolf Optimizer-Long Short-Term Memory

  • Jan 01, 2020
  • IEEE Access
  • Jinxin Pan +3
  • Research Article
  • Citations6

Spam text classification using LSTM Recurrent Neural Network

  • Sep 08, 2021
  • International Journal of Emerging Trends in Engineering Research
  • S Lai +27
  • Research Article
  • Citations14

Classification of acoustic emission sources produced by carbon/epoxy composite based on support vector machine

  • Jun 01, 2015
  • IOP Conference Series: Materials Science and Engineering
  • Peng Ding +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.