Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Yonghui Wu โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Yonghui Wu
58
papers ยท
7551
citations ยท
13
h-index
Beijing University of Technology ยท University of Science and Technology Beijing
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Conformer: Convolution-augmented Transformer for Speech Recognition
2020 ยท 2748 citations
Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis
2018 ยท 435 citations
W2v-BERT: Combining Contrastive Learning and Masked Language Modeling for Self-Supervised Speech Pre-Training
2021 ยท 262 citations
ContextNet: Improving Convolutional Neural Networks for Automatic Speech Recognition with Global Context
2020 ยท 257 citations
A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
2020 ยท 202 citations
Pushing the Limits of Semi-Supervised Learning for Automatic Speech Recognition
2020 ยท 200 citations
Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
2017 ยท 184 citations
Large-Scale Multilingual Speech Recognition with a Streaming End-to-End Model
2019 ยท 158 citations
Tacotron: Towards End-to-End Speech Synthesis
2017 ยท 152 citations
Towards Fast and Accurate Streaming End-to-End ASR
2020 ยท 113 citations
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 ยท 112 citations
Fully-hierarchical fine-grained prosody modeling for interpretable speech synthesis
2020 ยท 93 citations
A comparison of end-to-end models for long-form speech recognition
2019 ยท 85 citations
Non-Attentive Tacotron: Robust and Controllable Neural TTS Synthesis Including Unsupervised Duration Modeling
2020 ยท 73 citations
PnG BERT: Augmented BERT on Phonemes and Graphemes for Neural TTS
2021 ยท 61 citations
Top co-authors
Yu Zhang
ยท 26
Chung-Cheng Chiu
ยท 18
Ruoming Pang
ยท 16
Zhifeng Chen
ยท 15
Bo Li
ยท 11
Wei Han
ยท 11
Heiga Zen
ยท 10
James Qin
ยท 8
Jonathan Shen
ยท 8
Jiahui Yu
ยท 6
Yanzhang He
ยท 6
Anmol Gulati
ยท 5
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Generation
eess.AS
Voice Cloning
Audio Understanding
cs.CL
cs.SD
eess.SP