Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ming Lei — most-cited papers & profile · Speech Audio
← authors
·
overview
Ming Lei
21
papers ·
84
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep-FSMN for Large Vocabulary Continuous Speech Recognition
2018 · 13 citations
Streaming Chunk-Aware Multihead Attention for Online End-to-End Speech Recognition
2020 · 13 citations
Linear networks based speaker adaptation for speech synthesis
2018 · 12 citations
Universal ASR: Unifying Streaming and Non-Streaming ASR Using a Single Encoder-Decoder Model
2020 · 10 citations
Automatic Spelling Correction with Transformer for CTC-based End-to-End Speech Recognition
2019 · 9 citations
SAN-M: Memory Equipped Self-Attention for End-to-End Speech Recognition
2020 · 9 citations
Deep Feed-forward Sequential Memory Networks for Speech Synthesis
2018 · 8 citations
DeviceTTS: A Small-Footprint, Fast, Stable Network for On-Device Text-to-Speech
2020 · 8 citations
Simplified Self-Attention for Transformer-based End-to-End Speech Recognition
2020 · 1 citations
Speaker Embedding-aware Neural Diarization for Flexible Number of Speakers with Textual Information
2021 · 1 citations
Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs
2026
NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR
2026
Extremely Low Footprint End-to-End ASR System for Smart Device
2021
EMOVIE: A Mandarin Emotion Speech Dataset with a Simple Emotional Text-to-Speech Model
2021
BeamTransformer: Microphone Array-based Overlapping Speech Detection
2021
Top co-authors
Shiliang Zhang
· 11
Zhijie Yan
· 7
Zhifu Gao
· 4
Jie Gao
· 3
Yi Ren
· 3
Zhou Zhao
· 3
Guang Qiu
· 2
Haoneng Luo
· 2
Jiaqi Song
· 2
Lei Xie
· 2
Qian Chen
· 2
Xianliang Wang
· 2
Topics
Speech Recognition
Text-to-Speech
Speech Translation
eess.AS
cs.CL
cs.SD
Speaker Analysis
Voice Cloning
Audio Generation
Audio Understanding