• Home
  • Search
  • A neural fuzzy training approach for improving speech recognition
  • Cite Icon4
  • https://doi.org/10.1002/scj.4690240808Copy DOI Icon

A neural fuzzy training approach for improving speech recognition

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Abstract This paper proposes a new training method for the phoneme identification neural network called “neural fuzzy training.” In the proposed training, nondeterministic (fuzzy) class information is assigned to the training signal, in contrast to the traditional method where a deterministic class information is assigned.This study aims at the realization of a robust neural network, thereby improving the cumulative recognition rate of the phoneme identification and avoiding overtraining. The proposed neural fuzzy training is realized by backpropagation. In the conventional training, a deterministic phoneme class information is assigned to the training signal of the neural network as the value 1 or 0. However, in the proposed training, the fuzzy class information is assigned to the training signal for each training sample as the likelihood value between 0 and 1.In the proposed training method, the likelihood is calculated by the monotonically decreasing function (such as exp(−α · d2)) of the distance between the training sample and the closest sample belonging to each phoneme class. The proposed neural fuzzy training method has a problem in that a large amount of computation cost is required since the training signal is determined by calculating the distances to all training samples. To solve this problem, the representative samples in each phoneme class are defined and the likelihood to the phoneme classes are determined by calculating the distance between the representative sample and the training sample.By this simplification of the likelihood calculation, the computational cost to determine the training signal is reduced considerably. To demonstrate the usefulness of the neural fuzzy training, an experiment is conducted: /b, d, g, m, n, N/ identification, 18 consonant identification and phrase recognition using TDNN‐LR. The ATR database is used in the experiment. In the phoneme identification experiment, the speech samples which are extracted using the hand‐label is used. The TDNN is trained using speed samples uttered in word style, and the evaluation is performed using speech samples uttered in phrase style and in sentence style.In the phrase recognition experiment using TDNN‐LR, the TDNN is trained using speed samples uttered word style using a hand label. The evaluation is performed using speech samples uttered in phrase style. In either experiment, an improvement of using the fuzzy training can be observed. Especially, in the phrase recognition experiment using TDNN‐LR, the top recognition rate is improved from 71.2 percent to 80.9 percent, and the top 5th recognition rate is improved from 92.8 percent to 96.O percent. Furthermore, it appeared also that the neural fuzzy training is a high‐speed training method.

Similar Papers
  • Research Article
  • Citations23

A comparative study of interval and conventional training in thoroughbred racehorses.

  • Jun 01, 1990
  • Equine Veterinary Journal
  • J D Harkins +3
  • Research Article
  • Citations19

Generalization Effects of k-Neighbor Interpolation Training.

  • Sep 01, 1991
  • Neural computation
  • Takeshi Kawabata
  • Conference Article
  • Citations18

Partial Adversarial Training for Prediction Interval

  • Jul 01, 2018
  • H M Dipu Kabir +3
  • Conference Article

One-shot Training of Polynomial Cellular Neural Networks and applications in image processing

  • Jul 01, 2015
  • Antonio Arista-Jalife +1
  • Research Article
  • Citations12

LgNet: A Local-Global Network for Action Recognition and Beyond

  • Jan 01, 2023
  • IEEE Transactions on Multimedia
  • Jiaqi Zhou +4
  • PDF
  • Research Article
  • Citations11

Improving Photometric Redshift Estimates with Training Sample Augmentation

  • May 01, 2024
  • The Astrophysical Journal Letters
  • Irene Moskowitz +5
  • PDF
  • Research Article
  • Citations6

Performance comparison of neural network training methods based on wavelet packet transform for classification of five mental tasks

  • Jan 01, 2010
  • Journal of Biomedical Science and Engineering
  • Vijay Khare +3
  • Research Article
  • Citations8

A multi‐data training method for a deep neural network to improve the separation effect of simultaneous‐source data

  • Nov 14, 2022
  • Geophysical Prospecting
  • Kunxi Wang +3
  • Research Article
  • Citations2

Design the Training Program to Improve the Strength, Agility, and Quickness of the Table Tennis Players in Jing Zhou City

  • Sep 25, 2023
  • International Journal of Sociologies and Anthropologies Science Reviews
  • Chang Hu
  • Conference Article
  • Citations21

Hierarchical Transformer-Based Large-Context End-To-End ASR with Large-Context Knowledge Distillation

  • Jun 06, 2021
  • Ryo Masumura +5
  • PDF
  • Research Article

Algebraic Zero Error Training Method for Neural Networks Achieving Least Upper Bounds on Neurons and Layers

  • May 04, 2022
  • Computers
  • Juraj Kacur
  • Conference Article
  • Citations2

Finger vein recognition based on PCA and sparse representation

  • Sep 23, 2022
  • Lulu Zheng +1
  • Research Article
  • Citations5

Appearance-based representative samples refining method for palmprint recognition

  • Jul 06, 2012
  • Optical Engineering
  • Jiajun Wen
  • PDF
  • Research Article
  • Citations1

Identification, description and classification of consonants and vowel phonemes in Nambya language of Hwange district in Matabeleland North Province in Zimbabwe

  • Jan 30, 2024
  • International Journal of Science and Research Archive
  • Vincent Nyoni +1
  • Book Chapter
  • Citations2

Deterministic and Stochastic Logarithmic Barrier Function Methods for Neural Network Training

  • Jan 01, 1997
  • Theodore B Trafalis +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.