Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Rafael Valle — most-cited papers & profile · Speech Audio
← authors
·
overview
Rafael Valle
17
papers ·
98
citations ·
27
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Flowtron: an Autoregressive Flow-based Generative Network for Text-to-Speech Synthesis
2020 · 81 citations
Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
2024 · 8 citations
VANI: Very-lightweight Accent-controllable TTS for Native and Non-native speakers with Identity Preservation
2023 · 2 citations
Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
2025 · 1 citations
Generative Modeling for Low Dimensional Speech Attributes with Neural Spline Flows
2022 · 1 citations
Multilingual Multiaccented Multispeaker TTS with RADTTS
2023 · 1 citations
UniWav: Towards Unified Pre-training for Speech Representation Learning and Generation
2025
TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization
2024
A2SB: Audio-to-Audio Schrodinger Bridges
2025
One TTS Alignment To Rule Them All
2021
SelfVC: Voice Conversion With Iterative Refinement using Self Transformations
2023
Scaling NVIDIA's Multi-speaker Multi-lingual TTS Systems with Zero-Shot TTS to Indic Languages
2024
Top co-authors
Bryan Catanzaro
· 9
Rohan Badlani
· 6
Kevin J. Shih
· 4
Boris Ginsburg
· 3
Jo\~ao Felipe Santos
· 3
Paarth Neekhara
· 3
Shehzeen Hussain
· 3
Subhankar Ghosh
· 3
Akshit Arora
· 2
Jason Li
· 2
Sang-gil Lee
· 2
Zhifeng Kong
· 2
Topics
Audio Generation
Text-to-Speech
Voice Cloning
Speech Recognition
Music Generation
Speech Enhancement
Speech Translation
Multimodal Audio
Speaker Analysis