Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jinyu Li — most-cited papers & profile · Multimodal
← authors
·
overview
Jinyu Li
149
papers ·
1580
citations ·
0
h-index
Microsoft (United States) · Centre National de la Recherche Scientifique · Université Catholique de Lille · Université Sorbonne Nouvelle · Université de Lille · Microsoft Research (United Kingdom)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Continuous speech separation: dataset and analysis
2020 · 215 citations
Improving RNN Transducer Modeling for End-to-End Speech Recognition
2019 · 179 citations
Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
2023 · 163 citations
Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability
2020 · 97 citations
Advancing Acoustic-to-Word CTC Model
2018 · 90 citations
Adversarial Speaker Verification
2019 · 62 citations
Recent Advances in End-to-End Automatic Speech Recognition
2021 · 53 citations
Advancing Connectionist Temporal Classification With Attention Modeling
2018 · 49 citations
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition
2021 · 43 citations
Speaker Adaptation for Attention-Based End-to-End Speech Recognition
2019 · 41 citations
Adversarial Speaker Adaptation
2019 · 40 citations
Acoustic-To-Word Model Without OOV
2017 · 36 citations
Progressive Joint Modeling in Unsupervised Single-channel Overlapped Speech Recognition
2017 · 35 citations
Exploring Pre-training with Alignments for RNN Transducer based End-to-End Speech Recognition
2020 · 35 citations
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021 · 30 citations
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Speech Enhancement
Audio Generation
Speaker Analysis
Multimodal Audio
eess.AS
cs.CL