Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Pan Zhou — most-cited papers & profile · Speech Audio
← authors
·
overview
Pan Zhou
64
papers ·
340
citations ·
61
h-index
China University of Geosciences · Qingdao Institute of Marine Geology · Institute for Animal Reproduction · Huazhong University of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improving Generalization of Transformer for Speech Recognition with Parallel Schedule Sampling and Relative Positional Embedding
2019 · 23 citations
WNARS: WFST based Non-autoregressive Streaming End-to-End Speech Recognition
2021 · 9 citations
Exploring RNN-Transducer for Chinese Speech Recognition
2018 · 6 citations
An Online Attention-based Model for Speech Recognition
2018 · 6 citations
Adversarial Meta Sampling for Multilingual Low-Resource Speech Recognition
2020 · 3 citations
Modality Attention for End-to-End Audio-visual Speech Recognition
2018 · 1 citations
End-to-end contextual asr based on posterior distribution adaptation for hybrid ctc/attention system
2022 · 1 citations
MLCA-AVSR: Multi-Layer Cross Attention Fusion based Audio-Visual Speech Recognition
2024 · 1 citations
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
2024 · 1 citations
Towards Robust Overlapping Speech Detection: A Speaker-Aware Progressive Approach Using WavLM
2025
Wav-BERT: Cooperative Acoustic and Linguistic Representation Learning for Low-Resource Speech Recognition
2021
Automatic channel selection and spatial feature integration for multi-channel speech recognition across various array topologies
2023
Top co-authors
Wei Chen
· 7
Lei Xie
· 5
Pengcheng Guo
· 3
He Wang
· 2
Liang Lin
· 2
Xiaodan Liang
· 2
Ao Zhang
· 1
Dake Guo
· 1
Eng Siong Chng
· 1
Gang Liu
· 1
Jian Wu
· 1
Li Zhang
· 1
Topics
Speech Recognition
Speech Translation
Multimodal Audio
Audio Understanding
Text-to-Speech
Speaker Analysis