Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jay Mahadeokar — most-cited papers & profile · Speech Audio
← authors
·
overview
Jay Mahadeokar
28
papers ·
181
citations ·
14
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Transformer-Transducer: End-to-End Speech Recognition with Self-Attention
2019 · 66 citations
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
2023 · 45 citations
RNN-T For Latency Controlled ASR With Improved Beam Search
2019 · 34 citations
Dissecting User-Perceived Latency of On-Device E2E Speech Recognition
2021 · 26 citations
Contextualized Streaming End-to-End Speech Recognition with Trie-Based Deep Biasing and Shallow Fusion
2021 · 3 citations
Prompting Large Language Models with Speech Recognition Abilities
2023 · 2 citations
Improved Neural Language Model Fusion for Streaming Recurrent Neural Network Transducer
2020 · 1 citations
Deep Shallow Fusion for RNN-T Personalization
2020 · 1 citations
Streaming Transformer Transducer Based Speech Recognition Using Non-Causal Convolution
2021 · 1 citations
Multi-Head State Space Model for Speech Recognition
2023 · 1 citations
Towards Selection of Text-to-speech Data to Augment ASR Training
2023 · 1 citations
MELD: Mel-Spectrogram-Based Speech Language Modeling with Discrete Latent Variables
2026
Can Speech LLMs Think while Listening?
2025
CJST: CTC Compressor based Joint Speech and Text Training for Decoder-Only ASR
2024
Dynamic ASR Pathways: An Adaptive Masking Approach Towards Efficient Pruning of A Multilingual ASR Model
2023
Top co-authors
Ozlem Kalinli
· 21
Christian Fuegen
· 12
Chunyang Wu
· 11
Junteng Jia
· 11
Yuan Shangguan
· 11
Michael L. Seltzer
· 9
Yangyang Shi
· 8
Gil Keren
· 7
Duc Le
· 6
Mike Seltzer
· 6
Duc Le
· 5
Egor Lakomkin
· 5
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Multimodal Audio
Audio Generation
Speech Enhancement
Music Generation