Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Duc Le โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Duc Le
33
papers ยท
207
citations ยท
19
h-index
FPT University ยท Vietnam National University Ho Chi Minh City ยท Ho Chi Minh City University of Science ยท Ho Chi Minh City University of Technology ยท VNU University of Science
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
From Senones to Chenones: Tied Context-Dependent Graphemes for Hybrid Speech Recognition
2019 ยท 73 citations
Transformer-Transducer: End-to-End Speech Recognition with Self-Attention
2019 ยท 66 citations
Dissecting User-Perceived Latency of On-Device E2E Speech Recognition
2021 ยท 26 citations
Emformer: Efficient Memory Transformer Based Acoustic Model For Low Latency Streaming Speech Recognition
2020 ยท 13 citations
Weak-Attention Suppression For Transformer Based Speech Recognition
2020 ยท 5 citations
Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding
2021 ยท 5 citations
Evaluating User Perception of Speech Recognition System Quality with Semantic Distance Metric
2021 ยท 5 citations
Contextualized Streaming End-to-End Speech Recognition with Trie-Based Deep Biasing and Shallow Fusion
2021 ยท 3 citations
Neural-FST Class Language Model for End-to-End Speech Recognition
2022 ยท 3 citations
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
2024 ยท 2 citations
G2G: TTS-Driven Pronunciation Learning for Graphemic Hybrid ASR
2019 ยท 1 citations
Improved Neural Language Model Fusion for Streaming Recurrent Neural Network Transducer
2020 ยท 1 citations
Improving RNN Transducer Based ASR with Auxiliary Tasks
2020 ยท 1 citations
Deep Shallow Fusion for RNN-T Personalization
2020 ยท 1 citations
Learning ASR pathways: A sparse multilingual ASR model
2022 ยท 1 citations
Top co-authors
Michael L. Seltzer
ยท 18
Christian Fuegen
ยท 16
Ozlem Kalinli
ยท 15
Jay Mahadeokar
ยท 13
Suyoun Kim
ยท 9
Yangyang Shi
ยท 8
Ching-Feng Yeh
ยท 6
Chunyang Wu
ยท 5
Frank Zhang
ยท 5
Yuan Shangguan
ยท 5
Akshat Shrivastava
ยท 4
Abdelrahman Mohamed
ยท 3
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Multimodal Audio
Audio Generation
Music Generation
eess.AS
cs.CL