Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hao Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Hao Wang
267
papers ·
1290
citations ·
77
h-index
Stevens Institute of Technology · Xidian University · Norwegian University of Science and Technology · Hainan University · Shaanxi Normal University · Shanghai Maritime University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
2025 · 86 citations
Towards Natural and Controllable Cross-Lingual Voice Conversion Based on Neural TTS Model and Phonetic Posteriorgram
2021 · 13 citations
Speech2Slot: An End-to-End Knowledge-based Slot Filling from Speech
2021 · 4 citations
Towards Natural Bilingual and Code-Switched Speech Synthesis Based on Mix of Monolingual Recordings and Cross-Lingual Voice Conversion
2020 · 3 citations
Optimising Neural Speech Codecs for 300bps Communication using Reinforcement Learning
2026
ClariCodec: Optimising Neural Speech Codes for 200bps Communication using Reinforcement Learning
2026
Fun-ASR Technical Report
2025
CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training
2025
InspireMusic: Integrating Super Resolution and Large Language Model for High-Fidelity Long-Form Music Generation
2025
SPGM: Prioritizing Local Features for enhanced speech separation performance
2023
Top co-authors
Bin Ma
· 9
Chongjia Ni
· 7
Chong Zhang
· 5
Kun Zhou
· 5
Dianwen Ng
· 4
Wen Wang
· 4
Yukun Ma
· 4
Zhifu Gao
· 4
Zhihao Du
· 4
Changfeng Gao
· 3
Qian Chen
· 3
Xiang Lv
· 3
Topics
Audio Generation
Text-to-Speech
Speech Recognition
Voice Cloning
cs.SD
Speech Translation
Multimodal Audio
eess.AS
Audio Understanding
cs.CL