Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jinyu Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Jinyu Li
114
papers ·
1384
citations ·
0
h-index
Microsoft (United States) · Centre National de la Recherche Scientifique · Université Catholique de Lille · Université Sorbonne Nouvelle · Université de Lille · Microsoft Research (United Kingdom)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Continuous speech separation: dataset and analysis
2020 · 215 citations
Improving RNN Transducer Modeling for End-to-End Speech Recognition
2019 · 179 citations
Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
2023 · 163 citations
Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability
2020 · 97 citations
Advancing Acoustic-to-Word CTC Model
2018 · 90 citations
Recent Advances in End-to-End Automatic Speech Recognition
2021 · 53 citations
Advancing Connectionist Temporal Classification With Attention Modeling
2018 · 49 citations
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition
2021 · 43 citations
Acoustic-To-Word Model Without OOV
2017 · 36 citations
Exploring Pre-training with Alignments for RNN Transducer based End-to-End Speech Recognition
2020 · 35 citations
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021 · 30 citations
Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
2023 · 25 citations
Semantic Mask for Transformer based End-to-End Speech Recognition
2019 · 24 citations
High-Accuracy and Low-Latency Speech Recognition with Two-Head Contextual Layer Trajectory LSTM Model
2020 · 24 citations
Low Latency End-to-End Streaming Speech Recognition with a Scout Network
2020 · 22 citations
Top co-authors
Shujie Liu
· 43
Naoyuki Kanda
· 22
Long Zhou
· 20
Furu Wei
· 18
Takuya Yoshioka
· 18
Yashesh Gaur
· 18
Zhuo Chen
· 18
Yu Wu
· 16
Sanyuan Chen
· 15
Chengyi Wang
· 14
Zhong Meng
· 14
Jian Xue
· 13
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Generation
Speech Enhancement
Audio Understanding
Speaker Analysis
Multimodal Audio
Music Generation
Voice Cloning