Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hao Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Hao Li
218
papers ·
916
citations ·
8
h-index
Fudan University · Beijing Academy of Artificial Intelligence
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
On Modular Training of Neural Acoustics-to-Word Model for LVCSR
2018 · 35 citations
Using Optimal Ratio Mask as Training Target for Supervised Speech Separation
2017 · 22 citations
Modular End-to-end Automatic Speech Recognition Framework for Acoustic-to-word Model
2020 · 14 citations
DBNet: A Dual-branch Network Architecture Processing on Spectrum and Waveform for Single-channel Speech Enhancement
2021 · 11 citations
DeviceTTS: A Small-Footprint, Fast, Stable Network for On-Device Text-to-Speech
2020 · 8 citations
Multi-Channel Auto-Encoder for Speech Emotion Recognition
2018 · 5 citations
Integrated Speech Enhancement Method Based on Weighted Prediction Error and DNN for Dereverberation and Denoising
2017 · 2 citations
Speakerfilter-Pro: an improved target speaker extractor combines the time domain and frequency domain
2020 · 2 citations
EMPHASIS: An Emotional Phoneme-based Acoustic Model for Speech Synthesis System
2018 · 1 citations
STTATTS: Unified Speech-To-Text And Text-To-Speech Model
2024 · 1 citations
Multi-Loss Learning for Speech Emotion Recognition with Energy-Adaptive Mixup and Frame-Level Attention
2025
Data Augmentation for End-to-end Code-switching Speech Recognition
2020
Guided Training: A Simple Method for Single-channel Speaker Separation
2021
Minimally-Supervised Speech Synthesis with Conditional Diffusion Model and Language Model: A Comparative Study of Semantic Coding
2023
Learning Speech Representation From Contrastive Token-Acoustic Pretraining
2023
Top co-authors
Xueliang Zhang
· 7
Kai Yu
· 5
Guanglai Gao
· 2
Haoyu Li
· 2
Qi Liu
· 2
Tao Wang
· 2
Zhehuai Chen
· 2
Chenxing Li
· 1
Chunfeng Wang
· 1
Cong Wang
· 1
Hanan Aldarmaki
· 1
Hao Ni
· 1
Topics
Speech Recognition
Text-to-Speech
Audio Understanding
Speech Enhancement
Speech Translation
Audio Generation
Speaker Analysis
Multimodal Audio
Voice Cloning