Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Mu Yang — most-cited papers & profile · Speech Audio
← authors
·
overview
Mu Yang
13
papers ·
19
citations ·
46
h-index
Northwestern Polytechnical University · Tongji Hospital · Huazhong University of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Towards Lifelong Learning of Multilingual Text-To-Speech Synthesis
2021 · 2 citations
Bridging the Modality Gap: Softly Discretizing Audio Representation for LLM-based Automatic Speech Recognition
2025 · 1 citations
Improving Mispronunciation Detection with Wav2vec2-based Momentum Pseudo-Labeling for Accentedness and Intelligibility Assessment
2022 · 1 citations
Learning ASR pathways: A sparse multilingual ASR model
2022 · 1 citations
DiariST: Streaming Speech Translation with Speaker Diarization
2023 · 1 citations
Activation Steering for Accent-Neutralized Zero-Shot Text-To-Speech
2026
Emotion-Aware Prefix: Towards Explicit Emotion Control in Voice Conversion Models
2026
Spoken Language Intent Detection using Confusion2Vec
2019
Audiobox TTA-RAG: Improving Zero-Shot and Few-Shot Text-To-Audio with Retrieval-Augmented Generation
2024
Top co-authors
Andros Tjandra
· 2
Bowen Shi
· 1
David Zhang
· 1
Duc Le
· 1
Jian Xue
· 1
Jinyu Li
· 1
Ozlem Kalinli
· 1
Peidong Wang
· 1
Takuya Yoshioka
· 1
Tianlong Chen
· 1
Tong Wang
· 1
Xiaofei Wang
· 1
Topics
Speech Recognition
Audio Generation
Speech Translation
Multimodal Audio
Audio Understanding
eess.AS
Voice Cloning
Text-to-Speech
Speaker Analysis
Music Generation