Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ke Hu — most-cited papers & profile · Speech Audio
← authors
·
overview
Ke Hu
28
papers ·
486
citations ·
10
h-index
Tongji University · University of Alberta · Shenzhen University · Fudan University · Nanjing Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
2020 · 202 citations
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 · 112 citations
Deliberation Model Based Two-Pass End-to-End Speech Recognition
2020 · 76 citations
Transformer Based Deliberation for Two-Pass Speech Recognition
2021 · 24 citations
Massively Multilingual Shallow Fusion with Large Language Models
2023 · 14 citations
Phoneme-Based Contextualization for Cross-Lingual Speech Recognition in End-to-End Models
2019 · 7 citations
Textual Echo Cancellation
2020 · 7 citations
Improving Deliberation by Text-Only and Semi-Supervised Training
2022 · 7 citations
Multilingual and Fully Non-Autoregressive ASR with Large Language Model Fusion: A Comprehensive Study
2024 · 7 citations
Learning Word-Level Confidence For Subword End-to-End ASR
2021 · 2 citations
SALM-Duplex: Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
2025
VoiceTextBlender: Augmenting Large Language Models with Speech Capabilities via Single-Stage Joint Speech-Text Supervised Fine-Tuning
2024
Scaling Up Deliberation for Multilingual ASR
2022
A Deliberation-based Joint Acoustic and Text Decoder
2023
Mixture-of-Expert Conformer for Streaming Multilingual ASR
2023
Top co-authors
Bo Li
· 6
Yu Zhang
· 6
Zhehuai Chen
· 5
Boris Ginsburg
· 4
Ruoming Pang
· 3
Yanzhang He
· 3
Chung-Cheng Chiu
· 2
James Qin
· 2
Oleksii Hrinchuk
· 2
Shuo-yiin Chang
· 2
Vitaly Lavrukhin
· 2
Wei Li
· 2
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Multimodal Audio
Audio Understanding
Speech Enhancement