Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Sanyuan Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Sanyuan Chen
24
papers ·
252
citations ·
17
h-index
Fairchild Semiconductor (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
2023 · 163 citations
Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
2023 · 25 citations
SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data
2022 · 13 citations
Continuous Speech Separation with Conformer
2020 · 11 citations
UniSpeech-SAT: Universal Speech Representation Learning with Speaker Aware Pre-Training
2021 · 9 citations
Large-scale Self-Supervised Speech Representation Learning for Automatic Speaker Verification
2021 · 9 citations
VALL-E 2: Neural Codec Language Models are Human Parity Zero-Shot Text to Speech Synthesizers
2024 · 8 citations
Don't shoot butterfly with rifles: Multi-channel Continuous Speech Separation with Early Exit Transformer
2020 · 3 citations
TESSP: Text-Enhanced Self-Supervised Speech Pre-training
2022 · 3 citations
SpeechX: Neural Codec Language Model as a Versatile Speech Transformer
2023 · 3 citations
Self-Supervised Learning for speech recognition with Intermediate layer supervision
2021 · 2 citations
WavLLM: Towards Robust and Adaptive Speech Large Language Model
2024 · 2 citations
VALL-E R: Robust and Efficient Zero-Shot Text-to-Speech Synthesis via Monotonic Alignment
2024 · 1 citations
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling
2026
SAM Audio: Segment Anything in Audio
2025
Top co-authors
Shujie Liu
· 20
Jinyu Li
· 17
Yu Wu
· 14
Zhuo Chen
· 12
Furu Wei
· 11
Chengyi Wang
· 9
Long Zhou
· 9
Takuya Yoshioka
· 7
Jian Wu
· 6
Sheng Zhao
· 5
Yanqing Liu
· 5
Lingwei Meng
· 3
Topics
Speech Recognition
Speech Translation
Speech Enhancement
Text-to-Speech
Audio Generation
Audio Understanding
Speaker Analysis
Multimodal Audio
Voice Cloning
eess.AS