Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Dong Yu โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Dong Yu
178
papers ยท
1268
citations ยท
84
h-index
Seattle University ยท KLA (United States)
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Achieving Human Parity in Conversational Speech Recognition
2016 ยท 479 citations
DurIAN: Duration Informed Attention Network For Multimodal Synthesis
2019 ยท 94 citations
End-to-End Multi-Channel Speech Separation
2019 ยท 81 citations
FastDiff: A Fast Conditional Diffusion Model for High-Quality Speech Synthesis
2022 ยท 28 citations
BDDM: Bilateral Denoising Diffusion Models for Fast and High-Quality Speech Synthesis
2022 ยท 26 citations
Towards Robust Speaker Verification with Target Speaker Enhancement
2021 ยท 19 citations
DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs
2022 ยท 19 citations
Replay and Synthetic Speech Detection with Res2net Architecture
2020 ยท 18 citations
Minimum Bayes Risk Training of RNN-Transducer for End-to-End Speech Recognition
2019 ยท 16 citations
Deep Extractor Network for Target Speaker Recovery From Single Channel Speech Mixtures
2018 ยท 14 citations
Maximizing Mutual Information for Tacotron
2019 ยท 13 citations
DurIAN-SC: Duration Informed Attention Network based Singing Voice Conversion System
2020 ยท 13 citations
Recognizing Multi-talker Speech with Permutation Invariant Training
2017 ยท 12 citations
Speech-XLNet: Unsupervised Acoustic Model Pretraining For Self-Attention Networks
2019 ยท 11 citations
Learning Singing From Speech
2019 ยท 10 citations
Top co-authors
Dan Su
ยท 34
Meng Yu
ยท 34
Shi-Xiong Zhang
ยท 32
Yong Xu
ยท 25
Chunlei Zhang
ยท 18
Yiwen Shao
ยท 13
Chenxing Li
ยท 9
Helen Meng
ยท 9
Jun Wang
ยท 9
Hao Zhang
ยท 8
Jianwei Yu
ยท 8
Yuexian Zou
ยท 8
Topics
Speech Recognition
Speech Enhancement
Speech Translation
Audio Understanding
Audio Generation
Text-to-Speech
Speaker Analysis
Voice Cloning
Multimodal Audio
cs.SD