Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Rif A. Saurous — most-cited papers & profile · AI for Code
← authors
·
overview
Rif A. Saurous
21
papers ·
1325
citations ·
27
h-index
Google (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis
2018 · 475 citations
Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron
2018 · 205 citations
Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
2017 · 184 citations
Tacotron: Towards End-to-End Speech Synthesis
2017 · 152 citations
AutoMOS: Learning a non-intrusive assessor of naturalness-of-speech
2016 · 57 citations
VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking
2018 · 50 citations
Fixing a Broken ELBO
2017 · 44 citations
Uncovering Latent Style Factors for Expressive Speech Synthesis
2017 · 44 citations
Trainable Frontend For Robust and Far-Field Keyword Spotting
2016 · 12 citations
Exploring Tradeoffs in Models for Low-latency Speech Enhancement
2018 · 6 citations
Unsupervised Learning of Semantic Audio Representations
2017 · 5 citations
Multi-agent cooperation through in-context co-player inference
2026
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
2025
Embedded Universal Predictive Intelligence: a coherent framework for multi-agent learning
2025
On Using Backpropagation for Speech Texture Generation and Voice Conversion
2017
Topics
Text-to-Speech
Audio Generation
cs.AI
Speech Recognition
Speech Enhancement
stat.ML
cs.LG
cs.IT
eess.SP
Audio Understanding