• Home
  • Search
  • Automatic Voice Identification after Speech Resynthesis using PPG
  • Open Access IconOpen Access
  • https://doi.org/10.21437/odyssey.2024-27Copy DOI Icon

Automatic Voice Identification after Speech Resynthesis using PPG

  • Jun 18, 2024
  • Thibault Gaudier +3 more
Show More
  • Abstract
  • Literature Map
  • Similar Papers
Abstract

Speech resynthesis is a generic task for which we want to synthesize audio with another audio as input, which finds applications for media monitors and journalists. Among different tasks addressed by speech resynthesis, voice conversion preserves the linguistic information while modifying the identity of the speaker, and speech edition preserves the identity of the speaker but some words are modified. In both cases, we need to disentangle speaker and phonetic contents in intermediate representations. Phonetic PosteriorGrams (PPG) are a frame-level probabilistic representation of phonemes, and are usually considered speaker-independent. This paper presents a PPG-based speech resynthesis system. A perceptive evaluation assesses that it produces correct audio quality. Then, we demonstrate that an automatic speaker verification model is not able to recover the source speaker after re-synthesis with PPG, even when the model is trained on synthetic data.

Similar Papers
  • Research Article
  • Citations40

Who shall I say is calling? Validation of a caller recognition procedure in Bornean flanged male orangutan (Pongo pygmaeus wurmbii) long calls

  • Aug 05, 2016
  • Bioacoustics
  • Brigitte Spillmann +3
  • Conference Article
  • Citations8

Automatic Speaker Identification through Robust Time Domain Features and Hierarchical Classification Approach

  • May 12, 2018
  • Rashid Jahangir +4
  • Research Article
  • Citations2

Simultaneous speaker identification and watermarking

  • Jan 15, 2021
  • International Journal of Speech Technology
  • Basant S Abd El-Wahab +3
  • Research Article
  • Citations56

Microphone arrays and speaker identification

  • Jan 01, 1994
  • IEEE Transactions on Speech and Audio Processing
  • Qiguang Lin +2
  • Research Article
  • Citations8

Convolutive ICA-Based Forensic Speaker Identification Using Mel Frequency Cepstral Coefficients and Gaussian Mixture Models

  • Jul 02, 2013
  • The International Journal of Forensic Computer Science
  • Matheus Silveira +7
  • Book Chapter
  • Citations6

Emotional Speaker Identification by Humans and Machines

  • Jan 01, 2011
  • Yingchun Yang +2
  • Conference Article
  • Citations14

Bridging Mixture Density Networks with Meta-Learning for Automatic Speaker Identification

  • May 01, 2020
  • Ruirui Li +5
  • Research Article
  • Citations6

A two stage fuzzy decision classifier for speaker identification

  • Apr 01, 1996
  • Speech Communication
  • Pierre Castellano +1
  • Conference Article

Speaker identification employing waveform based speech CODEC

  • Aug 04, 2002
  • W.B Mikhael +1
  • Conference Article
  • Citations18

The optimization of perceptually-based features for speaker identification

  • May 23, 1989
  • L Xu +2
  • Research Article

Text Dependent Speaker Identification And Intruder Detection System

  • Jan 20, 2024
  • Revista Electronica de Veterinaria
  • N K Kaphungkui
  • Research Article
  • Citations1

Assessment of variation between and within speakers

  • Oct 01, 2003
  • The Journal of the Acoustical Society of America
  • Ruth Huntley Bahr
  • Conference Article
  • Citations13

Robust Analysis and Weighting on MFCC Components for Speech Recognition and Speaker Identification

  • Jul 01, 2007
  • Xi Zhou +4
  • Research Article
  • Citations10

Sensitivity of automatic speaker identification to SVD digital audio watermarking

  • Sep 01, 2015
  • International Journal of Speech Technology
  • Fathi E Abd El-Samie +6
  • Conference Article
  • Citations5

FASR: Effect of voice disguise

  • Oct 01, 2016
  • Kesiya Sebastian +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.