Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Edresson Casanova — most-cited papers & profile · Speech Audio
← authors
·
overview
Edresson Casanova
16
papers ·
72
citations ·
10
h-index
Nvidia (United Kingdom) · Nvidia (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone
2021 · 30 citations
CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese
2021 · 10 citations
BibleTTS: a large, high-fidelity, multilingual, and uniquely African speech corpus
2022 · 10 citations
CML-TTS A Multilingual Dataset for Speech Synthesis in Low-Resource Languages
2023 · 8 citations
Evaluation of Speech Representations for MOS prediction
2023 · 5 citations
ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion
2022 · 4 citations
XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model
2024 · 3 citations
Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
2025 · 1 citations
Brazilian Portuguese Speech Recognition Using Wav2vec 2.0
2021 · 1 citations
Tagarela - A Portuguese speech dataset from podcasts
2026
NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference
2025
HiFiTTS-2: A Large-Scale High Bandwidth Speech Dataset
2025
SALM-Duplex: Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
2025
FreeSVC: Towards Zero-shot Multilingual Singing Voice Conversion
2025
Speech2Phone: A Novel and Efficient Method for Training Speaker Recognition Models
2020
Top co-authors
Arnaldo Candido Junior
· 6
Jason Li
· 5
Anderson da Silva Soares
· 4
Frederico Santos de Oliveira
· 4
Lucas Rafael Stefanel Gris
· 4
Paarth Neekhara
· 4
Shehzeen Hussain
· 4
Subhankar Ghosh
· 4
Christopher Shulby
· 3
Ryan Langman
· 3
Xuesong Yang
· 3
Alef Iury Siqueira Ferreira
· 2
Topics
Audio Generation
Text-to-Speech
Speech Recognition
Speech Translation
Voice Cloning
Audio Understanding
Music Generation
Speaker Analysis
Multimodal Audio