Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Wei Zou — most-cited papers & profile · Speech Audio
← authors
·
overview
Wei Zou
31
papers ·
610
citations ·
32
h-index
Nanjing University of Information Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
2021 · 201 citations
Improving Transformer-based Speech Recognition Using Unsupervised Pre-training
2019 · 103 citations
Towards End-to-End Code-Switching Speech Recognition
2018 · 43 citations
Speech SIMCLR: Combining Contrastive and Reconstruction Objective for Self-supervised Speech Representation Learning
2020 · 14 citations
A Further Study of Unsupervised Pre-training for Transformer Based Speech Recognition
2020 · 9 citations
A comparable study of modeling units for end-to-end Mandarin speech recognition
2018 · 5 citations
DiDiSpeech: A Large Scale Mandarin Speech Corpus
2020 · 4 citations
Time Domain Adversarial Voice Conversion for ADD 2022
2022 · 4 citations
Advancing Speech Language Models by Scaling Supervised Fine-Tuning with Over 60,000 Hours of Synthetic Speech Dialogue Data
2024 · 1 citations
CLAR: CIF-Localized Alignment for Retrieval-Augmented Speech LLM-Based Contextual ASR
2026
LLaDA-TTS: Unifying Speech Synthesis and Zero-Shot Editing via Masked Diffusion Modeling
2026
Understanding the Modality Gap: An Empirical Study on the Speech-Text Alignment Mechanism of Large Speech Language Models
2025
Semantic Data Augmentation for End-to-End Mandarin Speech Recognition
2021
Top co-authors
Xiangang Li
· 10
Dongwei Jiang
· 6
Cheng Gong
· 1
Dan Su
· 1
Jianwei Sun
· 1
Junbo Zhang
· 1
Kun Han
· 1
Rui Yan
· 1
Sanjeev Khudanpur
· 1
Shinji Watanabe
· 1
Wei Wang
· 1
Xiaoyu Fan
· 1
Topics
Speech Recognition
Speech Translation
Text-to-Speech
cs.SD
Voice Cloning
Audio Generation
Audio Understanding
cs.CL
cs.AI
Speech Enhancement