Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jason Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Jason Li
30
papers ·
396
citations ·
0
h-index
Queen's University · Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Stochastic Gradient Methods with Layer-wise Adaptive Moments for Training of Deep Networks
2019 · 88 citations
Jasper: An End-to-End Convolutional Neural Acoustic Model
2019 · 44 citations
Mixed-Precision Training for NLP and Speech Recognition with OpenSeq2Seq
2018 · 41 citations
Training Neural Speech Recognition Systems with Synthetic Speech Augmentation
2018 · 41 citations
QuartzNet: Deep Automatic Speech Recognition with 1D Time-Channel Separable Convolutions
2019 · 31 citations
SpeakerNet: 1D Depth-wise Separable Convolutional Network for Text-Independent Speaker Recognition and Verification
2020 · 29 citations
Cross-Language Transfer Learning, Continuous Learning, and Domain Adaptation for End-to-End Automatic Speech Recognition
2020 · 20 citations
Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment
2024 · 8 citations
TTS-Transducer: End-to-End Speech Synthesis with Neural Transducer
2025 · 4 citations
Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
2025 · 1 citations
ACE-VC: Adaptive and Controllable Voice Conversion using Explicitly Disentangled Self-supervised Speech Representations
2023 · 1 citations
MagpieTTS-LF: Inference-Time Long-Form Speech Generation Without Training on Long-Form data
2026
Frame-Stacked Local Transformers For Efficient Multi-Codebook Speech Generation
2025
Align2Speak: Improving TTS for Low Resource Languages via ASR-Guided Online Preference Optimization
2025
NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference
2025
Top co-authors
Boris Ginsburg
· 13
Subhankar Ghosh
· 10
Vitaly Lavrukhin
· 8
Xuesong Yang
· 6
Oleksii Kuchaiev
· 5
Huyen Nguyen
· 3
Jocelyn Huang
· 3
Rafael Valle
· 3
Oleksii Hrinchuk
· 2
Yang Zhang
· 2
Zhehuai Chen
· 2
Bryan Catanzaro
· 1
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Speech Translation
Music Generation
eess.AS
Voice Cloning
Speaker Analysis
cs.AI
cs.CL