Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xinsheng Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Xinsheng Wang
16
papers ·
9
citations ·
16
h-index
Henan Institute of Science and Technology · Henan Polytechnic University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MsEmoTTS: Multi-scale emotion transfer, prediction, and control for emotional speech synthesis
2022 · 4 citations
Cross-speaker emotion disentangling and transfer for end-to-end speech synthesis
2021 · 3 citations
Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
2025 · 1 citations
Opencpop: A High-Quality Open Source Chinese Popular Song Corpus for Singing Voice Synthesis
2022 · 1 citations
DialoSpeech: Dual-Speaker Dialogue Generation with LLM and Flow Matching
2025
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
2025
Show and Speak: Directly Synthesize Spoken Description of Images
2020
Multi-speaker Multi-style Text-to-speech Synthesis With Single-speaker Single-style Training Data Scenarios
2021
Learn2Sing 2.0: Diffusion and Mutual Information-Based Target Speaker SVS by Learning from Singing Teacher
2022
AdaVITS: Tiny VITS for Low Computing Resource Speaker Adaptation
2022
Cross-speaker Emotion Transfer Based On Prosody Compensation for End-to-End Speech Synthesis
2022
Robust MelGAN: A robust universal neural vocoder for high-fidelity TTS
2022
Delivering Speaking Style in Low-resource Voice Conversion with Multi-factor Constraints
2022
UniSyn: An End-to-End Unified Model for Text-to-Speech and Singing Voice Synthesis
2022
StreamVoice: Streamable Context-Aware Language Modeling for Real-time Zero-Shot Voice Conversion
2024
Top co-authors
Lei Xie
· 11
Zhichao Wang
· 6
Qicong Xie
· 5
Yongmao Zhang
· 4
Heyang Xue
· 3
Yuanzhe Chen
· 3
Dan Su
· 2
Hanzhao Li
· 2
Jian Cong
· 2
Kun Song
· 2
Lei Xie
· 2
Mengxiao Bi
· 2
Topics
Audio Generation
Text-to-Speech
Voice Cloning
Speech Recognition
Music Generation
Speaker Analysis
Speech Enhancement
Speech Translation
Multimodal Audio