• Home
  • Search
  • Feature and model compensation for robust speech recognition
  • https://doi.org/10.1121/1.416494Copy DOI Icon

Feature and model compensation for robust speech recognition

Show More
  • Abstract
  • Literature Map
  • Similar Papers
Abstract

A mathematical framework based on maximum likelihood stochastic matching is proposed to perform feature and model compensation for robust speech recognition. Speech recognition is often formulated as a matching problem between the feature vectors extracted from a test utterance and a set of speech models or patterns obtained from some training corpra. It is well known that a speech recognizer often degrades in performance when the testing data are not acoustically similar to the training data. One way to improve is to find features that are invariant under all acoustic conditions and distortions. Some form of compensation is often required. The proposed stochastic matching approach assumes a structure or a form of the feature and/or model transformations. Together with a set of nuisance parameters, the transformations approximate the distortion in the test utterance. To decrease the acoustic mismatch between a test utterance and a given set of speech models, e.g., hidden Markov models, the stochastic matching algorithm estimates the nuisance parameters and then applies the feature/model transformations during speech recognition. Simple channel distortion can be approximated with linear transformations. For more complicated distortions, such as environmental, speaker, and combined mismatches, nonlinear compensating transformations are needed. These compensations give a significant performance improvement in speech recognition over the systems without them when utterances are affected by additive ambient noises and convolutional channel distortions.

Similar Papers
  • Research Article
  • Citations25

Combined speech enhancement and auditory modelling for robust distributed speech recognition

  • May 20, 2008
  • Speech Communication
  • Ronan Flynn +1
  • Research Article
  • Citations15

A study on model-based error rate estimation for automatic speech recognition

  • Nov 01, 2003
  • IEEE Transactions on Speech and Audio Processing
  • Chao-Shih Huang +2
  • Conference Article
  • Citations31

Adversarial Learning of Raw Speech Features for Domain Invariant Speech Recognition

  • Apr 01, 2018
  • Aditay Tripathi +3
  • Research Article
  • Citations16

Building a System for Arabic Dialects Identification based on Speech Recognition using Hidden Markov Models (HMMs)

  • Jul 10, 2021
  • DESIGN, CONSTRUCTION, MAINTENANCE
  • Zakaria Suliman Zubi +1
  • Research Article
  • Citations59

Audio-Visual Speech Recognition Using MPEG-4 Compliant Visual Features

  • Nov 28, 2002
  • EURASIP Journal on Advances in Signal Processing
  • Petar S Aleksic +3
  • Dissertation

Integrate template matching and statistical modeling for continuous speech recognition

  • Dec 01, 2011
  • Xie Sun
  • Research Article
  • Citations14

Intra- and Inter-frame Features for Automatic Speech Recognition

  • Jun 01, 2014
  • ETRI Journal
  • Sung Joo Lee +3
  • Research Article
  • Citations57

The Influence of Audibility on Speech Recognition With Nonlinear Frequency Compression for Children and Adults With Hearing Loss

  • Jul 01, 2014
  • Ear & Hearing
  • Ryan W Mccreery +5
  • Research Article
  • Citations23

Confusion analysis in phoneme based speech recognition in Hindi

  • Feb 01, 2020
  • Journal of Ambient Intelligence and Humanized Computing
  • Shobha Bhatt +2
  • Conference Article
  • Citations10

A theoretical analysis of speech recognition based on feature trajectory models

  • Oct 04, 2004
  • Yasuhiro Minami +3
  • Research Article

Research of Robust Feature for Speech Recognition

  • Jun 01, 2012
  • Advanced Materials Research
  • Xiang Hua Ren +1
  • Book Chapter
  • Citations10

Using Prosody in Fixed Stress Languages for Improvement of Speech Recognition

  • Mar 29, 2007
  • György Szaszák +1
  • Research Article
  • Citations8

Speech Recognition in Noise in Single-Sided Deaf Cochlear Implant Recipients Using Digital Remote Wireless Microphone Technology.

  • Jul 01, 2019
  • Journal of the American Academy of Audiology
  • Thomas Wesarg +8
  • Research Article
  • Citations1

Difference in Speech Recognition between a Default and Programmed Telecoil Program.

  • Jun 01, 2019
  • Journal of the American Academy of Audiology
  • Kimberly T Ledda +3
  • Book Chapter
  • Citations8

Speech Recognition Supported by Prosodic Information for Fixed Stress Languages

  • Sep 03, 2007
  • György Szaszák +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.