• Home
  • Search
  • Deep-Net: A Lightweight CNN-Based Speech Emotion Recognition System Using Deep Frequency Features.
  • Open Access IconOpen Access
  • Cite Icon163
  • https://doi.org/10.3390/s20185212Copy DOI Icon

Deep-Net: A Lightweight CNN-Based Speech Emotion Recognition System Using Deep Frequency Features.

  • Sep 12, 2020
  • Sensors
  • Tursunov Anvarjon +2 more
Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Artificial intelligence (AI) and machine learning (ML) are employed to make systems smarter. Today, the speech emotion recognition (SER) system evaluates the emotional state of the speaker by investigating his/her speech signal. Emotion recognition is a challenging task for a machine. In addition, making it smarter so that the emotions are efficiently recognized by AI is equally challenging. The speech signal is quite hard to examine using signal processing methods because it consists of different frequencies and features that vary according to emotions, such as anger, fear, sadness, happiness, boredom, disgust, and surprise. Even though different algorithms are being developed for the SER, the success rates are very low according to the languages, the emotions, and the databases. In this paper, we propose a new lightweight effective SER model that has a low computational complexity and a high recognition accuracy. The suggested method uses the convolutional neural network (CNN) approach to learn the deep frequency features by using a plain rectangular filter with a modified pooling strategy that have more discriminative power for the SER. The proposed CNN model was trained on the extracted frequency features from the speech data and was then tested to predict the emotions. The proposed SER model was evaluated over two benchmarks, which included the interactive emotional dyadic motion capture (IEMOCAP) and the berlin emotional speech database (EMO-DB) speech datasets, and it obtained 77.01% and 92.02% recognition results. The experimental results demonstrated that the proposed CNN-based SER system can achieve a better recognition performance than the state-of-the-art SER systems.

Loading PDF

Similar Papers
  • Research Article
  • Citations171

An ensemble 1D-CNN-LSTM-GRU model with data augmentation for speech emotion recognition

  • Feb 01, 2023
  • Expert Systems with Applications
  • Md Rayhan Ahmed +3
  • Research Article
  • Citations37

Unsupervised feature selection and NMF de-noising for robust Speech Emotion Recognition

  • Sep 24, 2020
  • Applied Acoustics
  • Surekha Reddy Bandela +1
  • Conference Article
  • Citations22

Designing and Evaluating Speech Emotion Recognition Systems: A Reality Check Case Study with IEMOCAP

  • Jun 04, 2023
  • Nikolaos Antoniou +3
  • Research Article
  • Citations4

SER: Performance Evaluation of CNN Model Along with an Overview of Available Indic Speech Datasets, and Transition of Classifiers From Traditional to Modern Era

  • Jun 26, 2023
  • ACM Transactions on Asian and Low-Resource Language Information Processing
  • Surbhi Khurana +2
  • Research Article

Comparative analysis of speech emotion recognition system using MLP, SVM, and CNN algorithms

  • Feb 10, 2026
  • Environmental Technology and Science Journal
  • B.A Omodunbi +3
  • Conference Article
  • Citations5

AI Based Web App and Framework for Detecting Emotions from Human Speech

  • Nov 11, 2022
  • Shubhaprasad Padhy +3
  • Book Chapter
  • Citations3

Speech Emotion Recognition Through Extraction of Various Emotional Features Using Machine and Deep Learning Classifiers

  • Jan 01, 2021
  • Kunal Jain +3
  • Research Article

Development of speech emotion recognition system using optimized convolutional neural network

  • Dec 19, 2024
  • LAUTECH Journal of Engineering and Technology
  • B F., Adebiyi +3
  • Conference Article
  • Citations2

Speech Emotion Recognition System for Arabic Speakers

  • Oct 22, 2022
  • Mai El Seknedy +1
  • Dissertation

Prosodic features of spontaneous emotional Singapore English speech

  • Jan 01, 2024
  • Rae Jia Xin Koh
  • Conference Article
  • Citations13

Speech emotion recognition with cross-lingual databases

  • Sep 14, 2014
  • Bo-Chang Chiou +1
  • Conference Article
  • Citations53

Speech Emotion Recognition using MFCC features and LSTM network

  • Sep 01, 2019
  • Harshawardhan S Kumbhar +1
  • PDF
  • Research Article
  • Citations19

Cross-Corpus Training Strategy for Speech Emotion Recognition Using Self-Supervised Representations

  • Aug 08, 2023
  • Applied Sciences
  • Miguel A Pastor +4
  • Research Article
  • Citations5

BiMER: Design and Implementation of a Bimodal Emotion Recognition System Enhanced by Data Augmentation Techniques

  • Jan 01, 2025
  • IEEE Access
  • Emrah Dikbiyik +2
  • PDF
  • Research Article
  • Citations62

Autoencoder With Emotion Embedding for Speech Emotion Recognition

  • Jan 01, 2021
  • IEEE Access
  • Chenghao Zhang +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.