Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shan Yang — most-cited papers & profile · Speech Audio
← authors
·
overview
Shan Yang
28
papers ·
88
citations ·
0
h-index
China Telecom (China) · China Telecom
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Data Efficient Voice Cloning from Noisy Samples with Domain Adversarial Training
2020 · 32 citations
Multi-band MelGAN: Faster Waveform Generation for High-Quality Text-to-Speech
2020 · 21 citations
Phonetic Posteriorgrams based Many-to-Many Singing Voice Conversion via Adversarial Training
2020 · 10 citations
EmoSteer-TTS: Fine-Grained and Training-Free Emotion-Controllable Text-to-Speech via Activation Steering
2025 · 9 citations
Statistical Parametric Speech Synthesis Using Generative Adversarial Networks Under A Multi-task Learning Framework
2017 · 6 citations
MsEmoTTS: Multi-scale emotion transfer, prediction, and control for emotional speech synthesis
2022 · 4 citations
Glow-WaveGAN: Learning Speech Representations from GAN-based Variational Auto-Encoder For High Fidelity Flow-based Speech Synthesis
2021 · 3 citations
DrawSpeech: Expressive Speech Synthesis Using Prosodic Sketches as Control Conditions
2025 · 2 citations
Learn2Sing: Target Speaker Singing Voice Synthesis by learning from a Singing Teacher
2020 · 1 citations
Covo-Audio Technical Report
2026
WavCube: Unifying Speech Representation for Understanding and Generation via Semantic-Acoustic Joint Modeling
2026
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
2025
Adversarial Feature Learning and Unsupervised Clustering based Speech Synthesis for Found Data with Acoustic and Textual Noise
2020
Exploiting Deep Sentential Context for Expressive End-to-End Speech Synthesis
2020
Accent and Speaker Disentanglement in Many-to-many Voice Conversion
2020
Top co-authors
Lei Xie
· 17
Dan Su
· 9
Dong Yu
· 5
Xinsheng Wang
· 3
Chenxing Li
· 2
Chunlei Zhang
· 2
Hanzhao Li
· 2
Liumeng Xue
· 2
Disong Wang
· 1
Guanglu Wan
· 1
Guanrou Yang
· 1
Guoqiao Yu
· 1
Topics
Audio Generation
Text-to-Speech
Voice Cloning
Speech Enhancement
Music Generation
Speech Recognition
eess.AS
cs.SD
cs.CL
cs.AI