Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Wenhao Guan — most-cited papers & profile · Speech Audio
← authors
·
overview
Wenhao Guan
15
papers ·
10
citations ·
5
h-index
Guilin University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model
2024 · 2 citations
Interpretable Style Transfer for Text-to-Speech with ControlVAE and Diffusion Bridge
2023 · 1 citations
MM-TTS: Multi-modal Prompt based Style Transfer for Expressive Text-to-Speech Synthesis
2023 · 1 citations
HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis
2026
SyncVoice: Towards Video Dubbing with Vision-Augmented Pretrained TTS Model
2025
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models
2025
ReFlow-VC: Zero-shot Voice Conversion Based on Rectified Flow and Speaker Feature Optimization
2025
Discl-VC: Disentangled Discrete Tokens and In-Context Learning for Controllable Zero-Shot Voice Conversion
2025
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec
2025
ReFlow-TTS: A Rectified Flow Model for High-fidelity Text-to-Speech
2023
LAFMA: A Latent Flow Matching Model for Text-to-Audio Generation
2024
Dynamic Language Group-Based MoE: Enhancing Code-Switching Speech Recognition with Hierarchical Routing
2024
Zero-Shot Sing Voice Conversion: built upon clustering-based phoneme representations
2024
Top co-authors
Qingyang Hong
· 10
Kaidi Wang
· 8
Hukai Huang
· 6
Lin Li
· 6
Peijie Chen
· 4
Yishuang Li
· 3
Jiayan Lin
· 2
Wangjin Zhou
· 2
Ziyue Jiang
· 2
Di Wu
· 1
Feng Dang
· 1
Feng Deng
· 1
Topics
Audio Generation
Text-to-Speech
Speech Recognition
Multimodal Audio
Voice Cloning
Speech Translation
Speech Enhancement
Music Generation
Audio Understanding
Speaker Analysis