Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Kainan Peng — most-cited papers & profile · Speech Audio
← authors
·
overview
Kainan Peng
12
papers ·
648
citations ·
13
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep Voice 2: Multi-Speaker Neural Text-to-Speech
2017 · 212 citations
Neural Voice Cloning with a Few Samples
2018 · 178 citations
Deep Voice 3: Scaling Text-to-Speech with Convolutional Sequence Learning
2017 · 103 citations
ClariNet: Parallel Wave Generation in End-to-End Text-to-Speech
2018 · 63 citations
WaveFlow: A Compact Flow-based Model for Raw Audio
2019 · 36 citations
Non-Autoregressive Neural Text-to-Speech
2019 · 26 citations
Multi-Speaker End-to-End Speech Synthesis
2019 · 25 citations
Incremental Text-to-Speech Synthesis with Prefix-to-Prefix Framework
2019 · 5 citations
SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment
2025
Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
2025
Zero-Shot Accent Conversion using Pseudo Siamese Disentanglement Network
2022
VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
2024
Top co-authors
Wei Ping
· 7
Mingbo Ma
· 4
Kexin Zhao
· 3
Zhenyu Tang
· 3
Andrew Gibiansky
· 2
Dongya Jia
· 2
Jiaxin Li
· 2
Jitong Chen
· 2
John Miller
· 2
Jonathan Raiman
· 2
Sercan O. Arik
· 2
Vimal Manohar
· 2
Topics
Audio Generation
Text-to-Speech
Voice Cloning
Speech Recognition
Speaker Analysis
Music Generation
Speech Translation
Speech Enhancement