Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiangang Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Xiangang Li
23
papers ·
859
citations ·
0
h-index
Alibaba Group (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep Speaker: an End-to-End Neural Speaker Embedding System
2017 · 427 citations
GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
2021 · 201 citations
Improving Transformer-based Speech Recognition Using Unsupervised Pre-training
2019 · 103 citations
Towards End-to-End Code-Switching Speech Recognition
2018 · 43 citations
Gram-CTC: Automatic Unit Selection and Target Decomposition for Sequence Labelling
2017 · 35 citations
Speech SIMCLR: Combining Contrastive and Reconstruction Objective for Self-supervised Speech Representation Learning
2020 · 14 citations
A Further Study of Unsupervised Pre-training for Transformer Based Speech Recognition
2020 · 9 citations
On Loss Functions and Recurrency Training for GAN-based Speech Enhancement Systems
2020 · 8 citations
Long Short-Term Memory based Convolutional Recurrent Neural Networks for Large Vocabulary Speech Recognition
2016 · 5 citations
A comparable study of modeling units for end-to-end Mandarin speech recognition
2018 · 5 citations
DiDiSpeech: A Large Scale Mandarin Speech Corpus
2020 · 4 citations
Time Domain Adversarial Voice Conversion for ADD 2022
2022 · 4 citations
Advancing Speech Language Models by Scaling Supervised Fine-Tuning with Over 60,000 Hours of Synthetic Speech Dialogue Data
2024 · 1 citations
FlowSE-GRPO: Training Flow Matching Speech Enhancement via Online Reinforcement Learning
2026
Fun-ASR Technical Report
2025
Top co-authors
Wei Zou
· 7
Dongwei Jiang
· 6
Shuaijiang Zhao
· 6
Ne Luo
· 4
Wubo Li
· 4
Tingwei Guo
· 3
Cheng Wen
· 2
Keyu An
· 2
Miao Cao
· 2
Ruixiong Zhang
· 2
Xiaoning Lei
· 2
Zhenyao Zhu
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Generation
Speech Enhancement
Speaker Analysis
Voice Cloning
Audio Understanding
Multimodal Audio