Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Yuxuan Wang โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Yuxuan Wang
37
papers ยท
1167
citations ยท
14
h-index
Inner Mongolia University ยท Advanced Energy Materials (United States)
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis
2018 ยท 475 citations
Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron
2018 ยท 205 citations
Tacotron: Towards End-to-End Speech Synthesis
2017 ยท 152 citations
Predicting Expressive Speaking Style From Text In End-To-End Speech Synthesis
2018 ยท 117 citations
Hierarchical Generative Modeling for Controllable Speech Synthesis
2018 ยท 45 citations
Uncovering Latent Style Factors for Expressive Speech Synthesis
2017 ยท 44 citations
VoiceFixer: Toward General Speech Restoration with Neural Vocoder
2021 ยท 25 citations
USTC-NELSLIP System Description for DIHARD-III Challenge
2021 ยท 20 citations
VoiceFixer: A Unified Framework for High-Fidelity Speech Restoration
2022 ยท 6 citations
Semi-Supervised Training for Improving Data Efficiency in End-to-End Speech Synthesis
2018 ยท 5 citations
Controllable and Lossless Non-Autoregressive End-to-End Text-to-Speech
2022 ยท 5 citations
PolyVoice: Language Models for Speech to Speech Translation
2023 ยท 5 citations
CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
2024 ยท 3 citations
A unified sequence-to-sequence front-end model for Mandarin text-to-speech synthesis
2019 ยท 2 citations
Speech enhancement with weakly labelled data from AudioSet
2021 ยท 2 citations
Top co-authors
Qiao Tian
ยท 6
RJ Skerry-Ryan
ยท 6
Yuping Wang
ยท 6
Daisy Stanton
ยท 5
Ming Tu
ยท 5
Qiuqiang Kong
ยท 5
Chuanzeng Huang
ยท 4
Qiao Tian
ยท 4
Rif A. Saurous
ยท 4
Eric Battenberg
ยท 3
Haohe Liu
ยท 3
Joel Shor
ยท 3
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Speech Enhancement
Audio Understanding
Speech Translation
Voice Cloning
Speaker Analysis
Multimodal Audio