Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Kaidi Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Kaidi Wang
10
papers ·
8
citations ·
3
h-index
Central South University · Sun Yat-sen University · Fudan University · Zhongshan Hospital · The First Affiliated Hospital, Sun Yat-sen University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model
2024 · 2 citations
HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis
2026
SyncVoice: Towards Video Dubbing with Vision-Augmented Pretrained TTS Model
2025
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models
2025
A Two-Stage Hierarchical Deep Filtering Framework for Real-Time Speech Enhancement
2025
ReFlow-VC: Zero-shot Voice Conversion Based on Rectified Flow and Speaker Feature Optimization
2025
Discl-VC: Disentangled Discrete Tokens and In-Context Learning for Controllable Zero-Shot Voice Conversion
2025
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec
2025
LAFMA: A Latent Flow Matching Model for Text-to-Audio Generation
2024
Top co-authors
Wenhao Guan
· 8
Qingyang Hong
· 7
Hukai Huang
· 4
Lin Li
· 4
Peijie Chen
· 4
Ziyue Jiang
· 2
Di Wu
· 1
Feng Dang
· 1
Feng Deng
· 1
Hongwu Ding
· 1
Hui Wang
· 1
Jian Luan
· 1
Topics
Audio Generation
Text-to-Speech
Multimodal Audio
Voice Cloning
Speech Recognition
Speech Enhancement
Speech Translation
Audio Understanding
Speaker Analysis
Music Generation