Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Rita Singh — most-cited papers & profile · Speech Audio
← authors
·
overview
Rita Singh
14
papers ·
37
citations ·
31
h-index
Max Healthcare · Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Pengi: An Audio Language Model for Audio Tasks
2023 · 20 citations
Voice Impersonation using Generative Adversarial Networks
2018 · 15 citations
On the Robust Approximation of ASR Metrics
2025 · 1 citations
DELULU: Discriminative Embedding Learning Using Latent Units for Speaker-Aware Self-Trained Speech Foundational Model
2025
OleSpeech-IV: A Large-Scale Multispeaker and Multilingual Conversational Speech Dataset with Diverse Topics
2025
Lost in Transcription, Found in Distribution Shift: Demystifying Hallucination in Speech Foundation Models
2025
BASS: Block-wise Adaptation for Speech Summarization
2023
Evaluating Speech Synthesis by Training Recognizers on Synthetic Speech
2023
Domain Adaptation for Contrastive Audio-Language Models
2024
SELM: Enhancing Speech Emotion Recognition for Out-of-Domain Scenarios
2024
Improving Speaker Representations Using Contrastive Losses on Multi-scale Features
2024
What Do Speech Foundation Models Not Learn About Speech?
2024
Top co-authors
Bhiksha Raj
· 11
Abdul Waheed
· 3
Hanin Atwany
· 3
Soham Deshmukh
· 3
Hira Dhamyal
· 2
Massa Baali
· 2
Roshan Sharma
· 2
Benjamin Elizalde
· 1
Chanwoo Kim
· 1
Chenfeng Miao
· 1
Dareen Alharthi
· 1
Dong Han
· 1
Topics
Speech Recognition
Audio Understanding
Multimodal Audio
Text-to-Speech
Speaker Analysis
Speech Translation
Audio Generation
cs.SD
cs.CL
Voice Cloning