• Home
  • Search
  • Complex Dynamic Neurons Improved Spiking Transformer Network for Efficient Automatic Speech Recognition
  • Cite Icon24
  • https://doi.org/10.1609/aaai.v37i1.25081Copy DOI Icon

Complex Dynamic Neurons Improved Spiking Transformer Network for Efficient Automatic Speech Recognition

  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

The spiking neural network (SNN) using leaky-integrated-and-fire (LIF) neurons has been commonly used in automatic speech recognition (ASR) tasks. However, the LIF neuron is still relatively simple compared to that in the biological brain. Further research on more types of neurons with different scales of neuronal dynamics is necessary. Here we introduce four types of neuronal dynamics to post-process the sequential patterns generated from the spiking transformer to get the complex dynamic neuron improved spiking transformer neural network (DyTr-SNN). We found that the DyTr-SNN could handle the non-toy automatic speech recognition task well, representing a lower phoneme error rate, lower computational cost, and higher robustness. These results indicate that the further cooperation of SNNs and neural dynamics at the neuron and network scales might have much in store for the future, especially on the ASR tasks.

Similar Papers
  • Conference Article
  • Citations1

Effect of Data Augmentation on DNN-Based VAD for Automatic Speech Recognition in Noisy Environment

  • Oct 13, 2020
  • Raufun Nahar +1
  • Conference Article
  • Citations12

Labeling unsegmented sequence data with DNN-HMM and its application for speech recognition

  • Sep 01, 2014
  • Xiangang Li +1
  • Research Article
  • Citations2

Automatic Speech Recognition Advancements for Indigenous Languages of the Americas

  • Jul 25, 2024
  • Applied Sciences
  • Monica Romero +2
  • Research Article
  • Citations10

Comparison of spiking neural networks with different topologies based on anti-disturbance ability under external noise

  • Feb 02, 2023
  • Neurocomputing
  • Lei Guo +3
  • Dissertation

SNN based ultra-low power system for voice assistant application

  • Jan 01, 2024
  • Chen Shen
  • Conference Article
  • Citations2

Train Your Classifier First: Cascade Neural Networks Training from Upper Layers to Lower Layers

  • Jun 06, 2021
  • Shucong Zhang +5
  • Conference Article
  • Citations3

Stabilising and Accelerating Light Gated Recurrent Units for Automatic Speech Recognition

  • Jun 04, 2023
  • Adel Moumen +1
  • Research Article
  • Citations58

Product of Gaussians for speech recognition

  • Jan 26, 2005
  • Computer Speech & Language
  • M.J.F Gales +1
  • Conference Article
  • Citations10

Early Stage LM Integration Using Local and Global Log-Linear Combination

  • Oct 25, 2020
  • Wilfried Michel +2
  • Research Article
  • Citations77

Noise-Robust Automatic Speech Recognition Using a Predictive Echo State Network

  • Jul 01, 2007
  • IEEE Transactions on Audio, Speech and Language Processing
  • Mark D Skowronski +1
  • Research Article
  • Citations6

Feature mapping using far-field microphones for distant speech recognition

  • Jul 19, 2016
  • Speech Communication
  • Ivan Himawan +3
  • Research Article
  • Citations66

Recognizing voice over IP: a robust front-end for speech recognition on the world wide web

  • Jun 01, 2001
  • IEEE Transactions on Multimedia
  • C Pelaez-Moreno +2
  • Conference Article
  • Citations1

Separate-to-Recognize: Joint Multi-target Speech Separation and Speech Recognition for Speaker-attributed ASR

  • Dec 11, 2022
  • Yuxiao Lin +5
  • Conference Article
  • Citations12

Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism

  • Sep 18, 2022
  • Kun Wei +2
  • Conference Article
  • Citations5

Entropy-based pruning of hidden units to reduce DNN parameters

  • Dec 01, 2016
  • Gautam Mantena +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.