Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuxuan Wang — most-cited papers & profile · Large Language Models
← authors
·
overview
Yuxuan Wang
37
papers ·
1167
citations ·
14
h-index
Inner Mongolia University · Advanced Energy Materials (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis
2018 · 475 citations
Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron
2018 · 205 citations
Tacotron: Towards End-to-End Speech Synthesis
2017 · 152 citations
Predicting Expressive Speaking Style From Text In End-To-End Speech Synthesis
2018 · 117 citations
Hierarchical Generative Modeling for Controllable Speech Synthesis
2018 · 45 citations
Hierarchical Generative Modeling for Controllable Speech Synthesis
2018 · 45 citations
Uncovering Latent Style Factors for Expressive Speech Synthesis
2017 · 44 citations
VoiceFixer: Toward General Speech Restoration with Neural Vocoder
2021 · 25 citations
USTC-NELSLIP System Description for DIHARD-III Challenge
2021 · 20 citations
VoiceFixer: A Unified Framework for High-Fidelity Speech Restoration
2022 · 6 citations
Semi-Supervised Training for Improving Data Efficiency in End-to-End Speech Synthesis
2018 · 5 citations
Controllable and Lossless Non-Autoregressive End-to-End Text-to-Speech
2022 · 5 citations
PolyVoice: Language Models for Speech to Speech Translation
2023 · 5 citations
CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
2024 · 3 citations
A unified sequence-to-sequence front-end model for Mandarin text-to-speech synthesis
2019 · 2 citations
Top co-authors
Daniel Chin
· 1
Gus Xia
· 1
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Speech Enhancement
Audio Understanding
Speech Translation
Voice Cloning
Speaker Analysis
Multimodal Audio
Vision-Language Models