Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shuai Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Shuai Wang
60
papers ·
130
citations ·
0
h-index
Qinghai University · Jinan University · Henan University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DiffRhythm: Blazingly Fast and Embarrassingly Simple End-to-End Full-Length Song Generation with Latent Diffusion
2025 · 51 citations
Voice activity detection in the wild: A data-driven approach using teacher-student training
2021 · 34 citations
SongBloom: Coherent Song Generation via Interleaved Autoregressive Sketching and Diffusion Refinement
2025 · 16 citations
BUT System for the Second DIHARD Speech Diarization Challenge
2020 · 5 citations
WEST: LLM based Speech Toolkit for Speech Understanding, Generation, and Interaction
2025 · 1 citations
WenetSpeech-Chuan: A Large-Scale Sichuanese Corpus with Rich Annotation for Dialectal Speech Processing
2025
Direct Preference Optimization for Speech Autoregressive Diffusion Models
2025
SenSE: Semantic-Aware High-Fidelity Universal Speech Enhancement
2025
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
2025
Traceable TTS: Toward Watermark-Free TTS with Strong Traceability
2025
Multi-Step Prediction and Control of Hierarchical Emotion Distribution in Text-to-Speech Synthesis
2025
Context-Aware Two-Step Training Scheme for Domain Invariant Speech Separation
2025
Hierarchical Emotion Prediction and Control in Text-to-Speech Synthesis
2024
Speech Separation with Pretrained Frontend to Minimize Domain Mismatch
2024
Top co-authors
Haizhou Li
· 4
Lei Xie
· 3
and Haizhou Li
· 2
Binbin Zhang
· 2
Johan Rohdin
· 2
Kun Zhou
· 2
Old\v{r}ich Plchot
· 2
Sho Inoue
· 2
Wupeng Wang
· 2
Zexu Pan
· 2
Zhao Guo
· 2
Anna Silnova
· 1
Topics
Audio Understanding
Text-to-Speech
Audio Generation
Speech Recognition
Speech Translation
Music Generation
Speech Enhancement
Speaker Analysis
Multimodal Audio
Voice Cloning