Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Yonghui Wu โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Yonghui Wu
34
papers ยท
5269
citations ยท
13
h-index
Beijing University of Technology ยท University of Science and Technology Beijing
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Conformer: Convolution-augmented Transformer for Speech Recognition
2020 ยท 2748 citations
Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis
2018 ยท 435 citations
W2v-BERT: Combining Contrastive Learning and Masked Language Modeling for Self-Supervised Speech Pre-Training
2021 ยท 262 citations
ContextNet: Improving Convolutional Neural Networks for Automatic Speech Recognition with Global Context
2020 ยท 257 citations
A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
2020 ยท 202 citations
Pushing the Limits of Semi-Supervised Learning for Automatic Speech Recognition
2020 ยท 200 citations
Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
2017 ยท 184 citations
Tacotron: Towards End-to-End Speech Synthesis
2017 ยท 152 citations
Towards Fast and Accurate Streaming End-to-End ASR
2020 ยท 113 citations
Fully-hierarchical fine-grained prosody modeling for interpretable speech synthesis
2020 ยท 93 citations
A comparison of end-to-end models for long-form speech recognition
2019 ยท 85 citations
Non-Attentive Tacotron: Robust and Controllable Neural TTS Synthesis Including Unsupervised Duration Modeling
2020 ยท 73 citations
PnG BERT: Augmented BERT on Phonemes and Graphemes for Neural TTS
2021 ยท 61 citations
Sequence-to-Sequence Models Can Directly Translate Foreign Speech
2017 ยท 55 citations
Hierarchical Generative Modeling for Controllable Speech Synthesis
2018 ยท 45 citations
Top co-authors
Ruoming Pang
ยท 15
Chung-Cheng Chiu
ยท 12
Yu Zhang
ยท 12
Ron J. Weiss
ยท 11
Ye Jia
ยท 10
Heiga Zen
ยท 9
Tara N. Sainath
ยท 8
Wei Han
ยท 8
Zhifeng Chen
ยท 8
Jonathan Shen
ยท 7
James Qin
ยท 6
Anmol Gulati
ยท 5
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Generation
Voice Cloning
Audio Understanding
Speaker Analysis
Speech Enhancement
Multimodal Audio