Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiaopeng Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Xiaopeng Wang
4
papers ·
0
citations ·
2
h-index
Kuaishou (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MM-Sonate: Multimodal Controllable Audio-Video Generation with Zero-Shot Voice Cloning
2026
M3-TTS: Multi-modal DiT Alignment & Mel-latent for Zero-shot High-fidelity Speech Synthesis
2025
DPI-TTS: Directional Patch Interaction for Fast-Converging and Style Temporal Modeling in Text-to-Speech
2024
Top co-authors
Chenxing Li
· 2
Chunyu Qiang
· 2
Ruibo Fu
· 2
Xuefei Liu
· 2
Yuankun Xie
· 2
Zhengqi Wen
· 2
Changsheng Li
· 1
Chen Zhang
· 1
Chunyu Qiang
· 1
Guanjun Li
· 1
Heng Xie
· 1
Jianhua Tao
· 1
Topics
Audio Generation
Text-to-Speech
Multimodal Audio
Voice Cloning
Music Generation