Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yifan Gong — most-cited papers & profile · Speech Audio
← authors
·
overview
Yifan Gong
21
papers ·
252
citations ·
39
h-index
Aerospace Center Hospital · Hubei University of Education · Universidad del Noreste
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Advancing Connectionist Temporal Classification With Attention Modeling
2018 · 49 citations
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition
2021 · 43 citations
Acoustic-To-Word Model Without OOV
2017 · 36 citations
On Addressing Practical Challenges for RNN-Transducer
2021 · 19 citations
PyKaldi2: Yet another speech toolkit based on Kaldi and PyTorch
2019 · 17 citations
Minimum Word Error Rate Training with Language Model Fusion for End-to-End Speech Recognition
2021 · 14 citations
RTMobile: Beyond Real-Time Mobile Acceleration of RNNs for Speech Recognition
2020 · 12 citations
Cracking the cocktail party problem by multi-beam deep attractor network
2018 · 3 citations
Speaker Separation Using Speaker Inventories and Estimated Speech
2020 · 2 citations
Minimum Latency Training Strategies for Streaming Sequence-to-Sequence ASR
2020 · 1 citations
Have best of both worlds: two-pass hybrid and E2E cascading framework for speech recognition
2021 · 1 citations
Internal Language Model Adaptation with Text-Only Data for End-to-End Speech Recognition
2021 · 1 citations
Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios
2025
Self-Teaching Networks
2019
Character-Aware Attention-Based End-to-End Speech Recognition
2020
Top co-authors
Jinyu Li
· 13
Liang Lu
· 5
Naoyuki Kanda
· 5
Yashesh Gaur
· 5
Eric Sun
· 4
Zhong Meng
· 4
Jian Xue
· 3
Xiong Xiao
· 3
Amit Das
· 2
Rui Zhao
· 2
Xie Chen
· 2
Alon Vinnikov
· 1
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Speech Enhancement
Speaker Analysis