Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Leda Sari — most-cited papers & profile · Speech Audio
← authors
·
overview
Leda Sari
9
papers ·
52
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
2023 · 45 citations
The Interspeech 2025 Speech Accessibility Project Challenge
2025 · 7 citations
CJST: CTC Compressor based Joint Speech and Text Training for Decoder-Only ASR
2024
Dynamic ASR Pathways: An Adaptive Masking Approach Towards Efficient Pruning of A Multilingual ASR Model
2023
Identify Speakers in Cocktail Parties with End-to-End Attention
2020
Biased Self-supervised learning for ASR
2022
Augmenting text for spoken language understanding with Large Language Models
2023
Frozen Large Language Models Can Perceive Paralinguistic Aspects of Speech
2024
Top co-authors
Jay Mahadeokar
· 4
Ozlem Kalinli
· 4
Junteng Jia
· 3
Chunyang Wu
· 2
Jinxi Guo
· 2
Ke Li
· 2
Mark Hasegawa-Johnson
· 2
Suyoun Kim
· 2
Wei Zhou
· 2
Abdelrahman Mohamed
· 1
Akshat Shrivastava
· 1
Andros Tjandra
· 1
Topics
Speech Recognition
Text-to-Speech
Audio Understanding
Multimodal Audio
Speaker Analysis
Audio Generation
Music Generation
Speech Translation