Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yu Wu — most-cited papers & profile · Speech Audio
← authors
·
overview
Yu Wu
33
papers ·
398
citations ·
0
h-index
Central South University · Xiangya Hospital Central South University · Guangzhou Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
2023 · 163 citations
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021 · 30 citations
Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
2023 · 25 citations
VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
2023 · 17 citations
SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data
2022 · 13 citations
Foundation Transformers
2022 · 13 citations
UniSpeech-SAT: Universal Speech Representation Learning with Speaker Aware Pre-Training
2021 · 9 citations
Large-scale Self-Supervised Speech Representation Learning for Automatic Speaker Verification
2021 · 9 citations
Wav2vec-Switch: Contrastive Learning from Original-noisy Speech Pairs for Robust Speech Recognition
2021 · 7 citations
Self-Supervised Learning for speech recognition with Intermediate layer supervision
2021 · 2 citations
Internal Language Model Adaptation with Text-Only Data for End-to-End Speech Recognition
2021 · 1 citations
LAMASSU: Streaming Language-Agnostic Multilingual Speech Recognition and Translation Using Neural Transducers
2022 · 1 citations
Investigation of Practical Aspects of Single Channel Speech Separation for ASR
2021
Streaming Speaker-Attributed ASR with Token-Level Speaker Embeddings
2022
Speech Pre-training with Acoustic Piece
2022
Top co-authors
Jinyu Li
· 16
Shujie Liu
· 15
Furu Wei
· 10
Sanyuan Chen
· 10
Chengyi Wang
· 9
Zhuo Chen
· 9
Long Zhou
· 7
Yao Qian
· 5
Jian Wu
· 4
Naoyuki Kanda
· 4
Takuya Yoshioka
· 4
Yashesh Gaur
· 4
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Speech Enhancement
Speaker Analysis
Audio Understanding
Multimodal Audio
Voice Cloning
Audio Generation
cs.ET