• Home
  • Search
  • Reconstructing Dual Learning for Neural Voice Conversion Using Relatively Few Samples
  • Cite Icon7
  • https://doi.org/10.1109/asru51503.2021.9687965Copy DOI Icon

Reconstructing Dual Learning for Neural Voice Conversion Using Relatively Few Samples

  • Dec 13, 2021
  • Aolan Sun +7 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

This paper introduces a dual learning system for neural voice conversion (DualVC) using relatively few samples based on the symmetry of the speech conversion task. The system contains a pair of sequence-to-sequence neural networks that have the same structure but are trained in opposite directions. The objective function of the dual model training is the sum of paired conversion loss and reconstruction loss during the dual training circle. The models in the two directions are trained alternately to guide each other by the corresponding reconstruction loss. Furthermore, curriculum learning techniques are used to load models in existing fields into the current task to accelerate the rapid iteration and convergence of the model. The experiment on the voice conversion task with the proposed DualVC and curriculum learning strategy obtained a comparable naturalness and similarity with only a 30% dataset than the BaseVC model trained on the full dataset.

Similar Papers
  • Conference Article
  • Citations2

Stargan-vc Based Cross-Domain Data Augmentation for Speaker Verification

  • Jun 04, 2023
  • Hang-Rui Hu +8
  • Preprint Article

CycleGAN Voice Conversion of Spectral Envelopes using Adversarial Weights

  • Oct 22, 2019
  • Zenodo (CERN European Organization for Nuclear Research)
  • Rafael Cunha Ferro +2
  • Conference Article
  • Citations50

Emotional voice conversion using deep neural networks with MCC and F0 features

  • Jun 01, 2016
  • Zhaojie Luo +2
  • Conference Article
  • Citations4

Enhancing Zero-Shot Many to Many Voice Conversion via Self-Attention VAE with Structurally Regularized Layers

  • Sep 01, 2022
  • Ziang Long +3
  • Conference Article

Voice conversion with pitch alteration using phase vocoder

  • Sep 01, 2017
  • Michal Lenarczyk +1
  • Conference Article

Voice Conversion System Based on Deep Neural Network Capable of Parallel Computation

  • Mar 01, 2018
  • Kunihiko Sato +1
  • Conference Article
  • Citations15

Perceptual-DualGAN: Perceptual Losses for Image to Image Translation with Generative Adversarial Nets

  • Jul 01, 2018
  • Xuexin Qu +4
  • Research Article
  • Citations1

An Improved Fully Convolutional Network Based on Post-Processing with Global Variance Equalization and Noise-Aware Training for Speech Enhancement

  • Jan 20, 2021
  • Journal of Advanced Computational Intelligence and Intelligent Informatics
  • Wenlong Li +3
  • Conference Article
  • Citations2

Speak Like a Dog: Human to Non-human creature Voice Conversion

  • Nov 07, 2022
  • Kohei Suzuki +3
  • Conference Article
  • Citations4

A massively parallel reverse modeling approach for semiconductor devices and circuits

  • Jan 01, 1997
  • Wu +1
  • Conference Article
  • Citations4

Learning Speech Representations with Flexible Hidden Feature Dimensions

  • Jun 04, 2023
  • Huaizhen Tang +4
  • Conference Article
  • Citations84

One-Shot Voice Conversion by Vector Quantization

  • May 01, 2020
  • Da-Yi Wu +1
  • Research Article
  • Citations79

Non-Parallel Training in Voice Conversion Using an Adaptive Restricted Boltzmann Machine

  • Nov 01, 2016
  • IEEE/ACM Transactions on Audio, Speech, and Language Processing
  • Toru Nakashika +2
  • Conference Article
  • Citations5

Voice Conversion of Tagalog Synthesized Speech Using Cycle-Generative Adversarial Networks (Cycle-GAN)

  • Oct 21, 2022
  • Jomari B Ganhinhin +3
  • Research Article

Development of mobile heart rate monitoring tools using artificial neural networks

  • Apr 04, 2024
  • Biomedical Radioelectronics
  • A.A Mikheev +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.