• Home
  • Search
  • First Steps Towards New Czech Voice Conversion System
  • Cite Icon3
  • https://doi.org/10.1007/11846406_48Copy DOI Icon

First Steps Towards New Czech Voice Conversion System

  • Jan 1, 2006
  • Zdeněk Hanzlíček +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Abstract In this paper we deal with initial experiments on creating a new Czech voice conversion system. Voice conversion (VC) is a process which modifies the speech signal produced by one (source) speaker so that it sounds like another (target) speaker. Using VC technique a new voice for speech synthesizer can be prepared with no need to record a huge amount of new speech data. The transformation is determined using equal sentences from both speakers; these sentences are time-aligned using modified dynamic time warping algorithm. The conversion is divided into two stages corresponding to the source-filter model of speech production. Within this work we employ conversion function based on Gaussian mixture model for transforming the spectral envelope described by line spectral frequencies. Residua are converted using so called residual prediction techniques. Unlike in other similar research works, we predict residua not from the transformed spectral envelope, but directly from the source speech. Four versions of residual prediction are described and compared in this study. Objective evaluation of converted speech using performance metrics shows that our system is comparable with similar existing VC systems.KeywordsGaussian Mixture ModelDynamic Time WarpingResidual SignalSpeech DataSource SpeechThese keywords were added by machine and not by the authors. This process is experimental and the keywords may be updated as the learning algorithm improves.

Similar Papers
  • Book Chapter
  • Citations4

Voice Conversion by Mapping the Spectral and Prosodic Features Using Support Vector Machine

  • Jan 01, 2009
  • Rabul Hussain Laskar +3
  • Research Article

A Voice Conversion System from Electrolarynx Speech to Preoperative Patient's Speech for Total Laryngectomy

  • Jan 01, 2026
  • OTO Open
  • Naoki Nishio +5
  • Conference Article
  • Citations1

A GMM based residual prediction method for voice conversion

  • Jan 01, 2005
  • Jing Xia +1
  • Research Article
  • Citations1

Speech naturalness improvement via $$\mathrm {\epsilon }$$ ϵ -closed extended vectors sets in voice conversion systems

  • Jan 12, 2017
  • Multidimensional Systems and Signal Processing
  • Mohammad Javad Jannati +2
  • Conference Article
  • Citations5

End-to-End Voice Conversion with Information Perturbation

  • Dec 11, 2022
  • Qicong Xie +4
  • Conference Article
  • Citations8

Improving the performance of HMM-based voice conversion using context clustering decision tree and appropriate regression matrix format

  • Sep 17, 2006
  • Long Qin +3
  • Research Article
  • Citations79

Non-Parallel Training in Voice Conversion Using an Adaptive Restricted Boltzmann Machine

  • Nov 01, 2016
  • IEEE/ACM Transactions on Audio, Speech, and Language Processing
  • Toru Nakashika +2
  • Conference Article
  • Citations3

Who is Speaking Actually? Robust and Versatile Speaker Traceability for Voice Conversion

  • Oct 26, 2023
  • Yanzhen Ren +5
  • Research Article
  • Citations1

포만트 공간에서의 주파수 변환을 이용한 이중 언어 음성 변환 연구

  • Dec 31, 2014
  • Phonetics and Speech Sciences
  • Yi-Geun Chae +3
  • Conference Article
  • Citations7

Voicy: Zero-Shot Non-Parallel Voice Conversion in Noisy Reverberant Environments

  • Aug 26, 2021
  • Alejandro Mottini +3
  • Conference Article
  • Citations11

Many-to-Many Voice Conversion Based Feature Disentanglement Using Variational Autoencoder

  • Aug 30, 2021
  • Manh Luong +1
  • Conference Article
  • Citations6

A Survey on Generative Adversarial Networks based Models for Many-to-many Non-parallel Voice Conversion

  • Mar 09, 2022
  • Yasmin Alaa +2
  • PDF
  • Research Article
  • Citations4

U2-VC: one-shot voice conversion using two-level nested U-structure

  • Nov 24, 2021
  • EURASIP Journal on Audio, Speech, and Music Processing
  • Fangkun Liu +4
  • Conference Article
  • Citations15

A dual alignment scheme for improved speech-to-singing voice conversion

  • Dec 01, 2017
  • Karthika Vijayan +2
  • Conference Article
  • Citations39

Singing Voice Conversion with Non-parallel Data

  • Mar 01, 2019
  • Xin Chen +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.