• Home
  • Search
  • GhostVec: Directly Extracting Speaker Embedding from End-to-End Speech Recognition Model Using Adversarial Examples
  • https://doi.org/10.1007/978-981-99-1645-0_40Copy DOI Icon

GhostVec: Directly Extracting Speaker Embedding from End-to-End Speech Recognition Model Using Adversarial Examples

  • Jan 1, 2023
  • Xiaojiao Chen +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Obtaining excellent speaker embedding representations can leverage the performance of a series of tasks, such as speaker/speech recognition, multi-speaker dialogue, and translation systems. The automatic speech recognition (ASR) system is trained with massive speech data and contains many speaker information. There are no existing attempts to protect the speaker embedding space of ASR from adversarial attacks. This paper proposes GhostVec, a novel method to export the speaker space from the ASR system without any external speaker verification system or real human voice as reference. More specifically, we extract speaker embedding from a transformer-based ASR system. Two kinds of targeted adversarial embedding (GhostVec) are proposed from features-level and embedding-level, respectively. The similarities are evaluated between GhostVecs and corresponding speakers randomly selected from Librispeech. Experiment results show that the proposed methods have superior performance in generating a similar embedding of the target speaker. We hope the preliminary discovery in this study to catalyze future downstream research speaker recognition-related topics.

Similar Papers
  • Research Article
  • Citations2

Adversarial Example Devastation and Detection on Speech Recognition System by Adding Random Noise

  • Jan 16, 2023
  • Journal of the Audio Engineering Society
  • Mingyu Dong +2
  • Research Article
  • Citations25

Combined speech enhancement and auditory modelling for robust distributed speech recognition

  • May 20, 2008
  • Speech Communication
  • Ronan Flynn +1
  • Research Article

Comparative study on noise-augmented training and its effect on adversarial robustness in ASR systems

  • Aug 26, 2025
  • Computer Speech & Language
  • Karla Pizzi +2
  • Research Article
  • Citations25

Enhancements in automatic Kannada speech recognition system by background noise elimination and alternate acoustic modelling

  • Jan 22, 2020
  • International Journal of Speech Technology
  • G Thimmaraja Yadava +1
  • Research Article

Efficient Detection of Targeted Adversarial Attacks in Automatic Speech Recognition Systems

  • Jan 01, 2025
  • IEEE Access
  • Daniyal Parveez +3
  • Conference Article
  • Citations36

Some insights from translating conversational telephone speech

  • May 01, 2014
  • Gaurav Kumar +3
  • PDF
  • Research Article
  • Citations2

Speech Recognition for the iCub Platform

  • Feb 12, 2018
  • Frontiers in Robotics and AI
  • Bertrand Higy +3
  • Research Article
  • Citations3

Challenges of Automatic Speech Recognition for medical interviews - research for Polish language

  • Jan 01, 2023
  • Procedia Computer Science
  • Karolina Kuligowska +2
  • Conference Article
  • Citations7

Development of a Low-Latency and Real-Time Automatic Speech Recognition System

  • Oct 13, 2020
  • Chee Siang Leow +3
  • Conference Article
  • Citations1

Accent neutralization for speech recognition of non-native speakers

  • Dec 02, 2019
  • Kacper Radzikowski +4
  • PDF
  • Research Article
  • Citations10

A Comparison of Hybrid and End-to-End ASR Systems for the IberSpeech-RTVE 2020 Speech-to-Text Transcription Challenge

  • Jan 17, 2022
  • Applied Sciences
  • Juan M Perero-Codosero +2
  • Conference Article
  • Citations15

Detecting Audio Attacks on ASR Systems with Dropout Uncertainty

  • Oct 25, 2020
  • Tejas Jayashankar +2
  • Preprint Article

Clinera ASR Benchmark: Evaluating Medical Code-Switching Automatic Speech Recognition for Arabic-English (ARZ-EN) (Preprint)

  • Mar 28, 2026
  • Ahmed Behairy +1
  • PDF
  • Research Article
  • Citations20

Dual supervised learning for non-native speech recognition

  • Jan 14, 2019
  • EURASIP Journal on Audio, Speech, and Music Processing
  • Kacper Radzikowski +3
  • PDF
  • Research Article
  • Citations33

Automatic Speech Recognition Predicts Speech Intelligibility and Comprehension for Listeners With Simulated Age-Related Hearing Loss.

  • Aug 29, 2017
  • Journal of Speech, Language, and Hearing Research
  • Lionel Fontan +8
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.