Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Furu Wei โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Furu Wei
103
papers ยท
1916
citations ยท
82
h-index
Microsoft (United States) ยท Microsoft (Finland) ยท Microsoft Research Asia (China) ยท Microsoft Research (United Kingdom)
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
2023 ยท 163 citations
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021 ยท 30 citations
Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
2023 ยท 25 citations
VibeVoice Technical Report
2025 ยท 23 citations
UniSpeech: Unified Speech Representation Learning with Labeled and Unlabeled Data
2021 ยท 20 citations
VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation
2023 ยท 17 citations
Pre-Training Transformer Decoder for End-to-End ASR Model with Unpaired Speech Data
2022 ยท 13 citations
SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data
2022 ยท 13 citations
Foundation Transformers
2022 ยท 13 citations
UniSpeech at scale: An Empirical Study of Pre-training Method on Large-Scale Speech Recognition Dataset
2021 ยท 9 citations
UniSpeech-SAT: Universal Speech Representation Learning with Speaker Aware Pre-Training
2021 ยท 9 citations
VALL-E 2: Neural Codec Language Models are Human Parity Zero-Shot Text to Speech Synthesizers
2024 ยท 8 citations
Boosting Large Language Model For Speech Synthesis: An Empirical Study
2023 ยท 7 citations
Neural Melody Composition from Lyrics
2018 ยท 4 citations
The YiTrans End-to-End Speech Translation System for IWSLT 2022 Offline Shared Task
2022 ยท 3 citations
Top co-authors
Shujie Liu
ยท 23
Jinyu Li
ยท 21
Long Zhou
ยท 17
Yu Wu
ยท 14
Sanyuan Chen
ยท 11
Chengyi Wang
ยท 9
Ziqiang Zhang
ยท 9
Zhuo Chen
ยท 7
Sheng Zhao
ยท 5
Yanqing Liu
ยท 5
Lirong Dai
ยท 4
Shaohan Huang
ยท 4
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Generation
Speech Enhancement
Audio Understanding
Multimodal Audio
Speaker Analysis
cs.AI
cs.SD