• Home
  • Search
  • Improving Deep Learning based Automatic Speech Recognition for Gujarati
  • Open Access IconOpen Access
  • Cite Icon15
  • https://doi.org/10.1145/3483446Copy DOI Icon

Improving Deep Learning based Automatic Speech Recognition for Gujarati

  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

We present a novel approach for improving the performance of an End-to-End speech recognition system for the Gujarati language. We follow a deep learning-based approach that includes Convolutional Neural Network, Bi-directional Long Short Term Memory layers, Dense layers, and Connectionist Temporal Classification as a loss function. To improve the performance of the system with the limited size of the dataset, we present a combined language model (Word-level language Model and Character-level language model)-based prefix decoding technique and Bidirectional Encoder Representations from Transformers-based post-processing technique. To gain key insights from our Automatic Speech Recognition (ASR) system, we used the inferences from the system and proposed different analysis methods. These insights help us in understanding and improving the ASR system as well as provide intuition into the language used for the ASR system. We have trained the model on the Microsoft Speech Corpus, and we observe a 5.87% decrease in Word Error Rate (WER) with respect to base-model WER.

Similar Papers
  • Research Article
  • Citations25

Enhancements in automatic Kannada speech recognition system by background noise elimination and alternate acoustic modelling

  • Jan 22, 2020
  • International Journal of Speech Technology
  • G Thimmaraja Yadava +1
  • Research Article
  • Citations2

Adversarial Example Devastation and Detection on Speech Recognition System by Adding Random Noise

  • Jan 16, 2023
  • Journal of the Audio Engineering Society
  • Mingyu Dong +2
  • Research Article
  • Citations4

End-to-end recognition of streaming Japanese speech using CTC and local attention

  • Jan 01, 2020
  • APSIPA Transactions on Signal and Information Processing
  • Jiahao Chen +2
  • Research Article
  • Citations25

Combined speech enhancement and auditory modelling for robust distributed speech recognition

  • May 20, 2008
  • Speech Communication
  • Ronan Flynn +1
  • Conference Article
  • Citations36

Some insights from translating conversational telephone speech

  • May 01, 2014
  • Gaurav Kumar +3
  • PDF
  • Research Article
  • Citations8

End-to-end automated speech recognition using a character based small scale transformer architecture

  • May 01, 2024
  • Expert Systems With Applications
  • Alexander Loubser +2
  • PDF
  • Research Article
  • Citations33

Automatic Speech Recognition Predicts Speech Intelligibility and Comprehension for Listeners With Simulated Age-Related Hearing Loss.

  • Aug 29, 2017
  • Journal of Speech, Language, and Hearing Research
  • Lionel Fontan +8
  • Book Chapter
  • Citations4

Gujarati Language Automatic Speech Recognition Using Integrated Feature Extraction and Hybrid Acoustic Model

  • Jan 01, 2023
  • Mohit Dua +1
  • PDF
  • Research Article
  • Citations2

Speech Recognition for the iCub Platform

  • Feb 12, 2018
  • Frontiers in Robotics and AI
  • Bertrand Higy +3
  • Research Article
  • Citations49

GFCC based discriminatively trained noise robust continuous ASR system for Hindi language

  • May 07, 2018
  • Journal of Ambient Intelligence and Humanized Computing
  • Mohit Dua +2
  • Conference Article
  • Citations1

Accent neutralization for speech recognition of non-native speakers

  • Dec 02, 2019
  • Kacper Radzikowski +4
  • Research Article
  • Citations3

Challenges of Automatic Speech Recognition for medical interviews - research for Polish language

  • Jan 01, 2023
  • Procedia Computer Science
  • Karolina Kuligowska +2
  • Conference Article
  • Citations7

Development of a Low-Latency and Real-Time Automatic Speech Recognition System

  • Oct 13, 2020
  • Chee Siang Leow +3
  • Research Article

Evaluating the Accuracy of Automatic Speech Recognition Systems in Home Healthcare Settings

  • Dec 01, 2025
  • Innovation in Aging
  • Dayoung Yu +7
  • Conference Article
  • Citations31

Contextual Language Model Adaptation for Conversational Agents

  • Sep 02, 2018
  • Anirudh Raju +7
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.