Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Carlos Busso — most-cited papers & profile · Speech Audio
← authors
·
overview
Carlos Busso
4
papers ·
6
citations ·
47
h-index
Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
End-to-end Audiovisual Speech Activity Detection with Bimodal Recurrent Neural Models
2018 · 3 citations
Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition
2024 · 3 citations
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
2025
Rethinking Speaker Embeddings for Speech Generation: Sub-Center Modeling for Capturing Intra-Speaker Diversity
2024
Top co-authors
Berrak Sisman
· 3
Ismail Rasim Ulgen
· 3
Ali N. Salman
· 1
Aurosweta Mahapatra
· 1
Fei Tao
· 1
John H. L. Hansen
· 1
Shreeram Suresh Chandra
· 1
Zongyang Du
· 1
Zongyang Du
· 1
Topics
Speech Recognition
Voice Cloning
Audio Generation
Speaker Analysis
Audio Understanding
Multimodal Audio
Music Generation
Speech Translation