Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jie Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Jie Zhang
150
papers ·
1514
citations ·
0
h-index
Queen's University Belfast · Kyung Hee University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Attentive Fusion Enhanced Audio-Visual Encoding for Transformer Based Robust Speech Recognition
2020 · 9 citations
Learning Contextually Fused Audio-visual Representations For Audio-visual Speech Recognition
2022 · 8 citations
Joint Training of Speech Enhancement and Self-supervised Model for Noise-robust ASR
2022 · 8 citations
Speech Enhancement Using Self-Supervised Pre-Trained Model and Vector Quantization
2022 · 5 citations
A Complementary Joint Training Approach Using Unpaired Speech and Text for Low-Resource Automatic Speech Recognition
2022 · 2 citations
Robust Data2vec: Noise-robust Speech Representation Learning for ASR by Combining Regression and Improved Contrastive Learning
2022 · 2 citations
A Composite Predictive-Generative Approach to Monaural Universal Speech Enhancement
2025 · 1 citations
VATLM: Visual-Audio-Text Pre-Training with Unified Masked Prediction for Speech Representation Learning
2022 · 1 citations
Semantic VAD: Low-Latency Voice Activity Detection for Speech Interaction
2023 · 1 citations
Rep2wav: Noise Robust text-to-speech Using self-supervised representations
2023 · 1 citations
CaSNet: Compress-and-Send Network Based Multi-Device Speech Enhancement Model for Distributed Microphone Arrays
2026
LABNet: A Lightweight Attentive Beamforming Network for Ad-hoc Multichannel Microphone Invariant Real-Time Speech Enhancement
2025
A Lightweight and Real-Time Binaural Speech Enhancement Model with Spatial Cues Preservation
2024
A Noise-Robust Self-supervised Pre-training Model Based Speech Representation Learning for Automatic Speech Recognition
2022
A Comparative Study on Multichannel Speaker-Attributed Automatic Speech Recognition in Multi-party Meetings
2022
Top co-authors
Lirong Dai
· 8
Yu Gu
· 6
Shihao Chen
· 5
Qian Chen
· 3
Shiliang Zhang
· 3
Fan Yu
· 2
Long Zhou
· 2
Yuchen Hu
· 2
Zhen-Hua Ling
· 2
Zhihao Du
· 2
Binxing Jiao
· 1
Daxin Jiang
· 1
Topics
Speech Recognition
Speech Enhancement
Text-to-Speech
Speech Translation
Audio Understanding
Audio Generation
Music Generation
Multimodal Audio
Speaker Analysis
cs.SD