Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Duc Le — most-cited papers & profile · Speech Audio
← authors
·
overview
Duc Le
12
papers ·
195
citations ·
19
h-index
FPT University · Vietnam National University Ho Chi Minh City · Ho Chi Minh City University of Science · Ho Chi Minh City University of Technology · VNU University of Science
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
From Senones to Chenones: Tied Context-Dependent Graphemes for Hybrid Speech Recognition
2019 · 73 citations
Transformer-Transducer: End-to-End Speech Recognition with Self-Attention
2019 · 66 citations
Dissecting User-Perceived Latency of On-Device E2E Speech Recognition
2021 · 26 citations
Emformer: Efficient Memory Transformer Based Acoustic Model For Low Latency Streaming Speech Recognition
2020 · 13 citations
Weak-Attention Suppression For Transformer Based Speech Recognition
2020 · 5 citations
Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding
2021 · 5 citations
Contextualized Streaming End-to-End Speech Recognition with Trie-Based Deep Biasing and Shallow Fusion
2021 · 3 citations
G2G: TTS-Driven Pronunciation Learning for Graphemic Hybrid ASR
2019 · 1 citations
Improved Neural Language Model Fusion for Streaming Recurrent Neural Network Transducer
2020 · 1 citations
Improving RNN Transducer Based ASR with Auxiliary Tasks
2020 · 1 citations
Deep Shallow Fusion for RNN-T Personalization
2020 · 1 citations
Dynamic Encoder Transducer: A Flexible Solution For Trading Off Accuracy For Latency
2021
Top co-authors
Michael L. Seltzer
· 10
Christian Fuegen
· 9
Jay Mahadeokar
· 6
Ching-Feng Yeh
· 5
Yangyang Shi
· 5
Chunyang Wu
· 4
Julian Chan
· 4
Ozlem Kalinli
· 4
Suyoun Kim
· 4
Yongqiang Wang
· 3
Yuan Shangguan
· 3
Frank Zhang
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding