• Home
  • Search
  • Confusability Measure Based Lexicon Optimization for Fast LVCSR Decoding
  • https://doi.org/10.14257/astl.2014.58.22Copy DOI Icon

Confusability Measure Based Lexicon Optimization for Fast LVCSR Decoding

Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

In this paper, we propose a lexicon optimization method based on confusability measure (CM) in order to reduce the decoding time for a large vo- cabulary continuous speech recognition (LVCSR) system. When lexicon is built or expanded for unseen words by using grapheme-to-phoneme (G2P) conver- sion, the lexicon size increases since G2P is generally realized by 1-to-N-best mapping. Thus, the proposed method prunes the confusable words in the lexi- con by a CM that is defined a linguistic distance between two phonemic se- quences. It is demonstrated from LVCSR experiments that the proposed lexicon optimization method achieves a relative real-time factor reduction of 23.13% on a task on the Wall Street Journal, compared to the 1-to-4-best G2P converted lexicon approach.

Similar Papers
  • Research Article
  • Citations22

Modelling Semantic Context of OOV Words in Large Vocabulary Continuous Speech Recognition

  • Feb 08, 2017
  • IEEE/ACM Transactions on Audio, Speech, and Language Processing
  • Imran Sheikh +3
  • Conference Article
  • Citations2

Large vocabulary continuous speech recognition based on cross-morpheme phonetic information

  • Oct 04, 2004
  • In-Jeong Choi +2
  • Conference Article

Large Vocabulary Continuous Audio-Visual Speech Recognition

  • Oct 02, 2018
  • George Sterpu
  • Research Article
  • Citations2

Confirmation Based Self-Learning Algorithm in LVCSR's Semi-supervised Incremental Learning

  • Jan 01, 2012
  • Procedia Engineering
  • Haifeng Li +2
  • Conference Article
  • Citations72

Combination of strongly and weakly constrained recognizers for reliable detection of OOVS

  • Mar 01, 2008
  • Lukas Burget +8
  • Conference Article
  • Citations8

Japanese large-vocabulary continuous-speech recognition using a business-newspaper corpus

  • Apr 21, 1997
  • T Matsuoka +5
  • Conference Article
  • Citations5

A Multi-Genre Urdu Broadcast Speech Recognition System

  • Nov 18, 2021
  • Erbaz Khan +3
  • Conference Article
  • Citations13

Recent improvements of the SpeeD Romanian LVCSR system

  • May 01, 2014
  • Horia Cucu +4
  • Conference Article
  • Citations4

An LVCSR Based Automatic Scoring Method in English Reading Tests

  • Aug 01, 2012
  • Junbo Zhang +2
  • Conference Article
  • Citations16

Malayalam Speech Recognition system and its application for visually impaired people

  • Dec 01, 2012
  • Anu V Anand +3
  • Research Article
  • Citations4

Construction and evaluation of language models based on stochastic context‐free grammar for speech recognition

  • Oct 23, 2002
  • Systems and Computers in Japan
  • Chiori Hori +3
  • Research Article
  • Citations1

A large-vocabulary continuous speech recognition system for Hindi

  • Sep 01, 2004
  • IBM Journal of Research and Development
  • Kumarm +2
  • Book Chapter
  • Citations25

Deep Neural Network Based Continuous Speech Recognition for Serbian Using the Kaldi Toolkit

  • Jan 01, 2015
  • Branislav Popović +4
  • Conference Article
  • Citations15

Speaker adaptation in the Philips system for large vocabulary continuous speech recognition

  • Apr 21, 1997
  • E Thelen +2
  • Research Article
  • Citations275

Fusion of Heterogeneous Speaker Recognition Systems in the STBU Submission for the NIST Speaker Recognition Evaluation 2006

  • Sep 01, 2007
  • IEEE Transactions on Audio, Speech, and Language Processing
  • Niko Brummer +9
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.