Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Wei Han โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Wei Han
51
papers ยท
4150
citations ยท
19
h-index
Nanjing University of Science and Technology ยท Guangzhou Railway Polytechnic
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Conformer: Convolution-augmented Transformer for Speech Recognition
2020 ยท 2748 citations
W2v-BERT: Combining Contrastive Learning and Masked Language Modeling for Self-Supervised Speech Pre-Training
2021 ยท 262 citations
ContextNet: Improving Convolutional Neural Networks for Automatic Speech Recognition with Global Context
2020 ยท 257 citations
Pushing the Limits of Semi-Supervised Learning for Automatic Speech Recognition
2020 ยท 200 citations
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 ยท 112 citations
A comparison of end-to-end models for long-form speech recognition
2019 ยท 85 citations
Noise2Music: Text-conditioned Music Generation with Diffusion Models
2023 ยท 50 citations
AudioPaLM: A Large Language Model That Can Speak and Listen
2023 ยท 41 citations
SLM: Bridge the thin gap between speech and text foundation models
2023 ยท 28 citations
Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling
2020 ยท 24 citations
Efficient Domain Adaptation for Speech Foundation Models
2023 ยท 19 citations
Exploring Targeted Universal Adversarial Perturbations to End-to-end ASR Models
2021 ยท 13 citations
RNN-T Models Fail to Generalize to Out-of-Domain Audio: Causes and Solutions
2020 ยท 12 citations
Accented Speech Recognition: Benchmarking, Pre-training, and Diverse Data
2022 ยท 12 citations
Supervised Contrastive Learning for Accented Speech Recognition
2021 ยท 9 citations
Top co-authors
Yu Zhang
ยท 18
Chung-Cheng Chiu
ยท 14
Ruoming Pang
ยท 11
Yonghui Wu
ยท 11
James Qin
ยท 8
Jiahui Yu
ยท 8
Anmol Gulati
ยท 6
Bo Li
ยท 6
Ankur Bapna
ยท 5
Izhak Shafran
ยท 4
Mingqiu Wang
ยท 4
Yuan Cao
ยท 4
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Speech Enhancement
Audio Generation
Multimodal Audio
Speaker Analysis
Music Generation