Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Kaizhi Qian — most-cited papers & profile · Speech Audio
← authors
·
overview
Kaizhi Qian
14
papers ·
294
citations ·
12
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
AUTOVC: Zero-Shot Voice Style Transfer with Only Autoencoder Loss
2019 · 195 citations
Unsupervised Speech Decomposition via Triple Information Bottleneck
2020 · 43 citations
ContentVec: An Improved Self-Supervised Speech Representation by Disentangling Speakers
2022 · 24 citations
PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition
2021 · 12 citations
Global Rhythm Style Transfer Without Text Transcriptions
2021 · 12 citations
Unsupervised Text-to-Speech Synthesis by Unsupervised Automatic Speech Recognition
2022 · 3 citations
Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Speech Processing
2022 · 2 citations
Deep Learning Based Speech Beamforming
2018 · 2 citations
Master-ASR: Achieving Multilingual Scalability and Low-Resource Adaptation in ASR with Modular Learning
2023 · 1 citations
Unlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction Tuning
2026
Rapverse: Coherent Vocals And Whole-body Motions Generations From Text
2024
On the Interplay Between Sparsity, Naturalness, Intelligibility, and Prosody in Speech Synthesis
2021
SpeechSplit 2.0: Unsupervised speech disentanglement for voice conversion Without tuning autoencoder Bottlenecks
2022
Towards Unsupervised Speech Recognition Without Pronunciation Models
2024
Top co-authors
Mark Hasegawa-Johnson
· 8
Shiyu Chang
· 8
David Cox
· 5
Yang Zhang
· 5
Heting Gao
· 3
Junrui Ni
· 3
Alexander H. Liu
· 2
Cheng-I Jeff Lai
· 2
Cheng-I Lai
· 2
James Glass
· 2
Xuesong Yang
· 2
Yang Zhang
· 2
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Speech Enhancement
Speech Translation
Audio Generation
Voice Cloning
Speaker Analysis
Multimodal Audio
Music Generation