Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Kaidi Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Kaidi Wang
14
papers ·
16
citations ·
3
h-index
Central South University · Sun Yat-sen University · Fudan University · Zhongshan Hospital · The First Affiliated Hospital, Sun Yat-sen University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SlimSpeech: Lightweight and Efficient Text-to-Speech with Slim Rectified Flow
2025 · 4 citations
Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model
2024 · 2 citations
SARA: A Dual-Stream VAE for High-Fidelity Speech Generation via Integrating Semantic and Acoustic Representations
2026
HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis
2026
SyncVoice: Towards Video Dubbing with Vision-Augmented Pretrained TTS Model
2025
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models
2025
A Two-Stage Hierarchical Deep Filtering Framework for Real-Time Speech Enhancement
2025
ReFlow-VC: Zero-shot Voice Conversion Based on Rectified Flow and Speaker Feature Optimization
2025
Discl-VC: Disentangled Discrete Tokens and In-Context Learning for Controllable Zero-Shot Voice Conversion
2025
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec
2025
LAFMA: A Latent Flow Matching Model for Text-to-Audio Generation
2024
Top co-authors
Lin Li
· 11
Qingyang Hong
· 11
Wenhao Guan
· 10
Peijie Chen
· 5
Hukai Huang
· 4
Weijie Wu
· 4
Shenghui Lu
· 2
Xie Chen
· 2
Ziyue Jiang
· 2
Daiyu Huang
· 1
di Wu
· 1
Feng Dang
· 1
Topics
Audio Generation
Text-to-Speech
Multimodal Audio
Speech Enhancement
cs.SD
Speech Translation
Voice Cloning
Speech Recognition
eess.AS
Speaker Analysis