Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Dong Yu โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Dong Yu
102
papers ยท
1284
citations ยท
84
h-index
Seattle University ยท KLA (United States)
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Achieving Human Parity in Conversational Speech Recognition
2016 ยท 479 citations
DurIAN: Duration Informed Attention Network For Multimodal Synthesis
2019 ยท 94 citations
End-to-End Multi-Channel Speech Separation
2019 ยท 81 citations
FastDiff: A Fast Conditional Diffusion Model for High-Quality Speech Synthesis
2022 ยท 28 citations
BDDM: Bilateral Denoising Diffusion Models for Fast and High-Quality Speech Synthesis
2022 ยท 26 citations
Towards Robust Speaker Verification with Target Speaker Enhancement
2021 ยท 19 citations
DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs
2022 ยท 19 citations
Replay and Synthetic Speech Detection with Res2net Architecture
2020 ยท 18 citations
The Microsoft 2016 Conversational Speech Recognition System
2016 ยท 16 citations
Minimum Bayes Risk Training of RNN-Transducer for End-to-End Speech Recognition
2019 ยท 16 citations
Deep Extractor Network for Target Speaker Recovery From Single Channel Speech Mixtures
2018 ยท 14 citations
Maximizing Mutual Information for Tacotron
2019 ยท 13 citations
DurIAN-SC: Duration Informed Attention Network based Singing Voice Conversion System
2020 ยท 13 citations
Recognizing Multi-talker Speech with Permutation Invariant Training
2017 ยท 12 citations
Speech-XLNet: Unsupervised Acoustic Model Pretraining For Self-Attention Networks
2019 ยท 11 citations
Top co-authors
Dan Su
ยท 31
Meng Yu
ยท 23
Chao Weng
ยท 18
Shi-Xiong Zhang
ยท 16
Shi-Xiong Zhang
ยท 13
Yong Xu
ยท 12
Helen Meng
ยท 9
Lianwu Chen
ยท 8
Chengzhu Yu
ยท 7
Chunlei Zhang
ยท 7
Max W. Y. Lam
ยท 7
Bo Wu
ยท 6
Topics
Speech Recognition
Speech Enhancement
Speech Translation
Audio Understanding
Audio Generation
Text-to-Speech
Speaker Analysis
Voice Cloning
Multimodal Audio
Music Generation