Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shi-Xiong Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Shi-Xiong Zhang
23
papers ·
48
citations ·
25
h-index
Capital One (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Audio-visual Recognition of Overlapped speech for the LRS2 dataset
2020 · 10 citations
Advancing Multi-talker ASR Performance with Large Language Models
2024 · 9 citations
End-to-End Attention based Text-Dependent Speaker Verification
2017 · 7 citations
NeuralEcho: A Self-Attentive Recurrent Neural Network For Unified Acoustic Echo Suppression And Speech Enhancement
2022 · 5 citations
Neural Spatio-Temporal Beamformer for Target Speech Separation
2020 · 3 citations
Self-supervised learning for audio-visual speaker diarization
2020 · 2 citations
Audio-visual Multi-channel Integration and Recognition of Overlapped Speech
2020 · 2 citations
Consistent Training and Decoding For End-to-end Speech Recognition Using Lattice-free MMI
2021 · 2 citations
EEND-SS: Joint End-to-End Neural Speaker Diarization and Speech Separation for Flexible Number of Speakers
2022 · 2 citations
Encrypted Speech Recognition using Deep Polynomial Networks
2019 · 1 citations
A comprehensive study of speech separation: spectrogram vs waveform separation
2019 · 1 citations
Audio-visual Multi-channel Recognition of Overlapped Speech
2020 · 1 citations
MIMO Self-attentive RNN Beamformer for Multi-speaker Speech Separation
2021 · 1 citations
Deep Neural Mel-Subband Beamformer for In-car Speech Separation
2022 · 1 citations
3D Neural Beamforming for Multi-channel Speech Separation Against Location Uncertainty
2023 · 1 citations
Top co-authors
Dong Yu
· 16
Meng Yu
· 8
Yong Xu
· 5
Yong Xu
· 5
Chunlei Zhang
· 4
Bo Wu
· 3
Chao Weng
· 3
Helen Meng
· 3
Jianwei Yu
· 3
Jian Wu
· 3
Rongzhi Gu
· 3
Xunying Liu
· 3
Topics
Speech Recognition
Speech Enhancement
Audio Understanding
Speech Translation
Speaker Analysis
Multimodal Audio
Text-to-Speech