Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Yu Wu โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Yu Wu
85
papers ยท
6290
citations ยท
36
h-index
Central South University ยท Wuhan University ยท Xiangya Hospital Central South University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
2023 ยท 163 citations
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021 ยท 30 citations
Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
2023 ยท 25 citations
Semantic Mask for Transformer based End-to-End Speech Recognition
2019 ยท 24 citations
Low Latency End-to-End Streaming Speech Recognition with a Scout Network
2020 ยท 22 citations
UniSpeech: Unified Speech Representation Learning with Labeled and Unlabeled Data
2021 ยท 20 citations
On the Comparison of Popular End-to-End Models for Large Scale Speech Recognition
2020 ยท 17 citations
VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
2023 ยท 17 citations
Curriculum Pre-training for End-to-End Speech Translation
2020 ยท 14 citations
Minimum Word Error Rate Training with Language Model Fusion for End-to-End Speech Recognition
2021 ยท 14 citations
SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data
2022 ยท 13 citations
Foundation Transformers
2022 ยท 13 citations
Continuous Speech Separation with Conformer
2020 ยท 11 citations
Developing Real-time Streaming Transformer Transducer for Speech Recognition on Large-scale Dataset
2020 ยท 9 citations
UniSpeech at scale: An Empirical Study of Pre-training Method on Large-Scale Speech Recognition Dataset
2021 ยท 9 citations
Top co-authors
Jinyu Li
ยท 32
Shujie Liu
ยท 31
Chengyi Wang
ยท 18
Zhuo Chen
ยท 16
Furu Wei
ยท 14
Sanyuan Chen
ยท 14
Jian Wu
ยท 10
Takuya Yoshioka
ยท 10
Long Zhou
ยท 9
Ming Zhou
ยท 6
Shuo Ren
ยท 5
Xie Chen
ยท 5
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Speech Enhancement
Audio Understanding
Speaker Analysis
Audio Generation
Multimodal Audio
Voice Cloning
Music Generation