• Home
  • Search
  • Discriminative Auditory-Based Featuresfor Robust Speech Recognition
  • Cite Icon27
  • https://doi.org/10.1109/tsa.2003.819951Copy DOI Icon

Discriminative Auditory-Based Featuresfor Robust Speech Recognition

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Recently, a new auditory-based feature extraction algorithm for robust speech recognition in noisy environments was proposed. The new features are derived by mimicking closely the human peripheral auditory process and the filters in the outer ear, middle ear, and inner ear are obtained from psychoacoustics literature with some manual adjustments. In this paper, we extend the auditory-based feature extraction algorithm and propose to further train the auditory-based filters through training. Using the data-driven approach, we optimize the filters by minimizing the subsequent recognition errors on a task. One significant contribution over similar efforts in the past (generally under the name of discriminative feature extraction) is that we make no assumption on the parametric form of the auditory-based filters. Instead, we only require the filters to be triangular-like: the filter weights have a maximum value in the middle and then monotonically decrease to both ends. Discriminative training of these constrained auditory-based filters leads to improved performance. Furthermore, we study the combined training procedure for both feature and acoustic model parameters. Our experiments show that the best performance can be obtained in a sequential procedure under the unified framework of MCE/GPD.

Similar Papers
  • Research Article
  • Citations20

A comparative study for Arabic speech recognition system in noisy environments

  • Apr 27, 2021
  • International Journal of Speech Technology
  • Abdelkbir Ouisaadane +1
  • Conference Article
  • Citations3

Mean normalization of power function based cepstral coefficients for robust speech recognition in noisy environment

  • May 01, 2014
  • Soonho Baek +1
  • Conference Article
  • Citations1

Compensating for noise and mismatch in speaker verification systems using approximate Bayesian inference

  • Mar 01, 2011
  • Ciira Wa Maina +1
  • Research Article
  • Citations107

Tissue-specific roles of Tbx1 in the development of the outer, middle and inner ear, defective in 22q11DS patients

  • Apr 06, 2006
  • Human Molecular Genetics
  • Jelena S Arnold +6
  • PDF
  • Research Article

Development of Modeling Methods for the Outer and Middle Ear of a Human

  • Dec 27, 2019
  • Microsystems, Electronics and Acoustics
  • Anastasiia Olehivna Harasiuk +1
  • Conference Article
  • Citations16

Hands free continuous speech recognition in noisy environment using a four microphone array

  • May 09, 1995
  • D Giuliani +3
  • Conference Article
  • Citations1

On the impact of normalizing power-based features on robustness against noise for speech recognition

  • Oct 01, 2016
  • Hilman F Pardede
  • Research Article

Improving speech intelligibility for hearing aid users using a speech segregation algorithm.

  • Oct 01, 2010
  • The Journal of the Acoustical Society of America
  • Srikanth Vishnubhotla +1
  • Research Article

Enhancing Robustness of Arabic Speech Recognition in Noisy Environments Using Advanced Feature Extraction and Denoising Techniques Based on Deep Learning Models

  • Dec 04, 2025
  • Circuits, Systems, and Signal Processing
  • Lallouani Bouchakour +2
  • Research Article
  • Citations70

Robust speech recognition in noisy environments based on subband spectral centroid histograms

  • Mar 01, 2006
  • IEEE Transactions on Audio, Speech and Language Processing
  • B Gajic +1
  • Research Article
  • Citations21

DISCRIMINATIVE LEARNING OF VISUAL DATA FOR AUDIOVISUAL SPEECH RECOGNITION

  • Mar 01, 1999
  • International Journal on Artificial Intelligence Tools
  • Alexandrina Rogozan
  • Research Article
  • Citations20

Enhancing Computer Digital Signal Processing through the Utilization of RNN Sequence Algorithms

  • Dec 30, 2023
  • International Journal of Computer Science and Information Technology
  • Hongjie Niu +4
  • Research Article
  • Citations1

An Approach Towards to Real Time AI Desktop Voice Assistant

  • May 26, 2024
  • International Journal of Advanced Research in Science, Communication and Technology
  • Dr Anup Bhange +4
  • Research Article
  • Citations18

Feature compensation based on switching linear dynamic model

  • Jun 01, 2005
  • IEEE Signal Processing Letters
  • Nam Soo Kim +2
  • Book Chapter
  • Citations1

Combining Multi-band and Frequency-Filtering Techniques for Speech Recognition in Noisy Environments

  • Jan 01, 2000
  • Peter Jančovič +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.