The Neteasegames System For Voice Conversion Challenge 2020 With Vector-quantization Variational Autoencoder And Wavenet
2020 Β· Haitong Zhang
Abstract
This paper presents the description of our submitted system for Voice Conversion Challenge (VCC) 2020 with vector-quantization variational autoencoder (VQ-VAE) with WaveNet as the decoder, i.e., VQ-VAE-WaveNet. VQ-VAE-WaveNet is a nonparallel VAE-based voice conversion that reconstructs the acoustic features along with separating the linguistic information with speaker identity. The model is further improved with the WaveNet cycle as the decoder to generate the high-quality speech waveform, since WaveNet, as an autoregressive neural vocoder, has achieved the SoTA result of waveform generation. In practice, our system can be developed with VCC 2020 dataset for both Task 1 (intra-lingual) and Task 2 (cross-lingual). However, we only submit our system for the intra-lingual voice conversion task. The results of VCC 2020 demonstrate that our system VQ-VAE-WaveNet achieves: 3.04 mean opinion score (MOS) in naturalness and a 3.28 average score in similarity ( the speaker similarity percentage
Authors
(none)
Tags
Stats
Related papers
- Baseline System Of Voice Conversion Challenge 2020 With Cyclic Variational Autoencoder And Parallel Wavegan (2020)4.24
- The NU Voice Conversion System For The Voice Conversion Challenge 2020: On The Effectiveness Of Sequence-to-sequence Models And Autoregressive Neural Vocoders (2020)3.58
- The Academia Sinica Systems Of Voice Conversion For VCC2020 (2020)3.58
- VQVC+: One-shot Voice Conversion By Vector Quantization And U-net Architecture (2020)13.34
- Refined Wavenet Vocoder For Variational Autoencoder Based Voice Conversion (2018)7.50
- Unsupervised Acoustic Unit Representation Learning For Voice Conversion Using Wavenet Auto-encoders (2020)7.16
- Voice Conversion Challenge 2020: Intra-lingual Semi-parallel And Cross-lingual Voice Conversion (2020)12.74
- Efficient Non-autoregressive GAN Voice Conversion Using Vqwav2vec Features And Dynamic Convolution (2022)0.00