• Home
  • Search
  • Many-to-Many Voice Conversion Based Feature Disentanglement Using Variational Autoencoder
  • Cite Icon11
  • https://doi.org/10.21437/interspeech.2021-2086Copy DOI Icon

Many-to-Many Voice Conversion Based Feature Disentanglement Using Variational Autoencoder

  • Aug 30, 2021
  • Manh Luong +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Voice conversion is a challenging task which transforms the voice characteristics of a source speaker to a target speaker without changing linguistic content. Recently, there have been many works on many-to-many Voice Conversion (VC) based on Variational Autoencoder (VAEs) achieving good results, however, these methods lack the ability to disentangle speaker identity and linguistic content to achieve good performance on unseen speaker scenarios. In this paper, we propose a new method based on feature disentanglement to tackle many to many voice conversion. The method has the capability to disentangle speaker identity and linguistic content from utterances, it can convert from many source speakers to many target speakers with a single autoencoder network. Moreover, it naturally deals with the unseen target speaker scenarios. We perform both objective and subjective evaluations to show the competitive performance of our proposed method compared with other state-of-the-art models in terms of naturalness and target speaker similarity.

Similar Papers
  • Conference Article
  • Citations4

Enhancing Zero-Shot Many to Many Voice Conversion via Self-Attention VAE with Structurally Regularized Layers

  • Sep 01, 2022
  • Ziang Long +3
  • Conference Article
  • Citations5

End-to-End Voice Conversion with Information Perturbation

  • Dec 11, 2022
  • Qicong Xie +4
  • PDF
  • Research Article
  • Citations4

U2-VC: one-shot voice conversion using two-level nested U-structure

  • Nov 24, 2021
  • EURASIP Journal on Audio, Speech, and Music Processing
  • Fangkun Liu +4
  • Research Article
  • Citations1

A Multidomain Generative Adversarial Network for Hoarse-to-Normal Voice Conversion

  • Oct 01, 2023
  • Journal of voice : official journal of the Voice Foundation
  • Minghang Chu +7
  • Conference Article
  • Citations3

Arabic speech transformation using MFCC in GMM

  • Jul 01, 2012
  • Rania Elmanfaloty +2
  • Conference Article
  • Citations14

Investigation of Using Disentangled and Interpretable Representations for One-shot Cross-lingual Voice Conversion

  • Sep 02, 2018
  • Seyed Hamidreza Mohammadi +1
  • Conference Article
  • Citations3

DVQVC: An Unsupervised Zero-Shot Voice Conversion Framework

  • Jun 04, 2023
  • Dayong Li +2
  • Conference Article
  • Citations66

High-order sequence modeling using speaker-dependent recurrent temporal restricted boltzmann machines for voice conversion

  • Sep 14, 2014
  • Toru Nakashika +2
  • Research Article
  • Citations163

Speaker Transformation Algorithm using Segmental Codebooks (STASC)

  • Jul 01, 1999
  • Speech Communication
  • Levent M Arslan
  • Research Article
  • Citations8

Voice conversion with SI-DNN and KL divergence based mapping without parallel training data

  • Nov 30, 2018
  • Speech Communication
  • Feng-Long Xie +2
  • Research Article
  • Citations1

Speech naturalness improvement via $$\mathrm {\epsilon }$$ ϵ -closed extended vectors sets in voice conversion systems

  • Jan 12, 2017
  • Multidimensional Systems and Signal Processing
  • Mohammad Javad Jannati +2
  • Book Chapter
  • Citations4

Voice Conversion by Mapping the Spectral and Prosodic Features Using Support Vector Machine

  • Jan 01, 2009
  • Rabul Hussain Laskar +3
  • Conference Article
  • Citations1

A GMM based residual prediction method for voice conversion

  • Jan 01, 2005
  • Jing Xia +1
  • Conference Article
  • Citations4

Personalizing TTS Voices for Progressive Dysarthria

  • Jul 27, 2021
  • Yunxin Zhao +3
  • Research Article
  • Citations28

Voice Conversion for Persons with Amyotrophic Lateral Sclerosis.

  • Dec 25, 2019
  • IEEE Journal of Biomedical and Health Informatics
  • Yunxin Zhao +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.