Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
RJ Skerry-Ryan — most-cited papers & profile · Speech Audio
← authors
·
overview
RJ Skerry-Ryan
18
papers ·
1287
citations ·
16
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis
2018 · 475 citations
Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron
2018 · 205 citations
Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
2017 · 184 citations
Tacotron: Towards End-to-End Speech Synthesis
2017 · 152 citations
Predicting Expressive Speaking Style From Text In End-To-End Speech Synthesis
2018 · 117 citations
Uncovering Latent Style Factors for Expressive Speech Synthesis
2017 · 44 citations
Effective Use of Variational Embedding Capacity in Expressive End-to-End Speech Synthesis
2019 · 44 citations
Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning
2019 · 26 citations
Semi-Supervised Generative Modeling for Controllable Speech Synthesis
2019 · 15 citations
Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling
2021 · 11 citations
Semi-Supervised Training for Improving Data Efficiency in End-to-End Speech Synthesis
2018 · 5 citations
Speaker Generation
2021 · 1 citations
Robust and Unbounded Length Generalization in Autoregressive Transformer-Based Text-to-Speech
2024
Location-Relative Attention Mechanisms For Robust Long-Form Speech Synthesis
2019
Spoken Question Answering and Speech Continuation Using Spectrogram-Powered LLM
2023
Top co-authors
Daisy Stanton
· 10
Eric Battenberg
· 8
Soroosh Mariooryad
· 6
Yuxuan Wang
· 6
Matt Shannon
· 5
Rif A. Saurous
· 5
Ron J. Weiss
· 4
Tom Bagby
· 4
Yonghui Wu
· 4
David Kao
· 3
Joel Shor
· 3
Julian Salazar
· 3
Topics
Text-to-Speech
Audio Generation
Voice Cloning
Speech Enhancement
Speaker Analysis
Music Generation
Speech Recognition
Speech Translation
Audio Understanding
Multimodal Audio