VCTK corpus
Emerging5papers using it
2021first seen
The VCTK corpus is a dataset that contains recordings of English speech from multiple speakers, used to evaluate the performance of audio processing systems, such as codecs.
Papers using VCTK corpus (5)
- AudioDec: An Open-source Streaming High-fidelity Neural Audio CodecMDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High
Sampling Rate and Low Bitrate ScenariosConditional Deep Hierarchical Variational Autoencoder for Voice
ConversionmdctGAN: Taming transformer-based GAN for speech super-resolution with
Modified DCT spectraWho is Authentic Speaker