• Home
  • Search
  • Robust Front-End Based on MVA and HEQ Post-processing for Arabic Speech Recognition Using Hidden Markov Model Toolkit (HTK)
  • Cite Icon5
  • https://doi.org/10.1109/aiccsa.2017.180Copy DOI Icon

Robust Front-End Based on MVA and HEQ Post-processing for Arabic Speech Recognition Using Hidden Markov Model Toolkit (HTK)

  • Oct 1, 2017
  • Elhem Techini +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

This paper describes a study of a set of features based on cepstral mean and variance normalization (CMVN) plus auto regressive moving average (ARMA) filtering technique which is called MVA and on histogram equalization (HEQ) for robust speech recognition. First, we use MVA then HEQ in combination with CMVN and ARMA filtering as a post-processing module to mel frequency cepstral coefficients (MFCC), Relative Spectral-Perceptual linear prediction (RASTA-PLP) and power normalized cepstral coefficients (PNCC) features to improve the performance of the automatic speech recognition (ASR) system. The results on the Arabic database task have shown that both methods MVA and HEQ+ARMA improves the success rate for all features compared to the baseline system however HEQ was not found to perform better than MVA. The results also provide that RASTA-PLP outperforms PNCC and MFCC features.

Similar Papers
  • Book Chapter
  • Citations4

Detection of Operation Type and Order for Digital Speech

  • Dec 22, 2019
  • Tingting Wu +3
  • Conference Article
  • Citations3

Robust language identification using Power Normalized Cepstral Coefficients

  • Aug 01, 2015
  • Arup Kumar Dutta +1
  • Conference Article

Robustifying cepstral features by mitigating the outlier effect for noisy speech recognition

  • Jul 01, 2013
  • Hao-Teng Fan +3
  • Book Chapter
  • Citations20

A Comparative Study of Feature and Score Normalization for Speaker Verification

  • Jan 01, 2005
  • Rong Zheng +2
  • Research Article
  • Citations63

Optimization of temporal filters for constructing robust features in speech recognition

  • May 01, 2006
  • IEEE Transactions on Audio, Speech and Language Processing
  • Jeih-Weih Hung +1
  • Research Article
  • Citations46

Significance of analytic phase of speech signals in speaker verification

  • Feb 26, 2016
  • Speech Communication
  • Karthika Vijayan +2
  • Conference Article
  • Citations7

Particle Swarm Optimisation of Mel-frequency Cepstral Coefficients computation for the classification of asphyxiated infant cry

  • Oct 01, 2010
  • A Zabidi +4
  • Conference Article
  • Citations5

Efficient MFCC feature extraction on graphics processing units

  • Jan 01, 2013
  • Haofeng Kou +3
  • Conference Article
  • Citations27

Robust speaker recognition based on improved GFCC

  • Oct 01, 2016
  • Xiaoyuan Shi +2
  • Conference Article
  • Citations8

Weighted cosine distance features for speaker verification

  • Dec 01, 2015
  • C Santhosh Kumar +3
  • Research Article
  • Citations7

Real-time prediction of upcoming respiratory events via machine learning using snoring sound signal.

  • Apr 12, 2021
  • Journal of Clinical Sleep Medicine
  • Bochun Wang +6
  • Conference Article
  • Citations5

Integration of articulatory knowledge and voicing features based on DNN/HMM for Mandarin speech recognition

  • Jul 01, 2015
  • Ying-Wei Tan +3
  • Conference Article
  • Citations1

Input Fusion of MFCC and SCMC Features for Acoustic Scene Classification using DNN

  • Dec 01, 2018
  • Chandrasekhar Paseddula +1
  • Research Article
  • Citations70

Automatic speech recognition with an adaptation model motivated by auditory processing

  • Jan 01, 2006
  • IEEE Transactions on Audio, Speech and Language Processing
  • M Holmberg +2
  • Conference Article
  • Citations21

Incorporating frequency masking filtering in a standard MFCC feature extraction algorithm

  • Jan 01, 2004
  • Weizhong Zhu +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.