Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuhang Dai — most-cited papers & profile · Speech Audio
← authors
·
overview
Yuhang Dai
6
papers ·
12
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Unveiling the Potential of LLM-Based ASR on Chinese Open-Source Datasets
2024 · 12 citations
SoulX-Duplug: Plug-and-Play Streaming State Prediction Module for Realtime Full-Duplex Speech Conversation
2026
SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis
2026
X-Talk: On the Underestimated Potential of Modular Speech-to-Speech Dialogue System
2025
WenetSpeech-Yue: A Large-scale Cantonese Speech Corpus with Multi-dimensional Annotation
2025
AISHELL-5: The First Open-Source In-Car Multi-Channel Multi-Speaker Speech Dataset for Automatic Speech Diarization and Recognition
2025
Top co-authors
Hanke Xie
· 2
Hanlin Wen
· 2
Haopeng Lin
· 2
He Wang
· 2
Hongfei Xue
· 2
Hui Bu
· 2
Jiale Qian
· 2
and Lei Xie
· 1
Binbin Zhang
· 1
Bingshen Mu
· 1
Chaochao Lu
· 1
Chengyou Wang
· 1
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Audio Generation
Music Generation
Speech Enhancement
Speaker Analysis
Speech Translation
Multimodal Audio