Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Qiushi Zhu — most-cited papers & profile · Speech Audio
← authors
·
overview
Qiushi Zhu
11
papers ·
12
citations ·
10
h-index
University of Science and Technology of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Speech Enhancement Using Self-Supervised Pre-Trained Model and Vector Quantization
2022 · 5 citations
Wav2code: Restore Clean Speech Representations via Codebook Lookup for Noise-Robust ASR
2023 · 2 citations
Cross-Modal Global Interaction and Local Alignment for Audio-Visual Speech Recognition
2023 · 2 citations
VATLM: Visual-Audio-Text Pre-Training with Unified Masked Prediction for Speech Representation Learning
2022 · 1 citations
Rep2wav: Noise Robust text-to-speech Using self-supervised representations
2023 · 1 citations
Gradient Remedy for Multi-Task Learning in End-to-End Noise-Robust Speech Recognition
2023
Noise-aware Speech Enhancement using Diffusion Probabilistic Model
2023
Multichannel AV-wav2vec2: A Framework for Learning Multichannel Multi-Modal Speech Representation
2024
Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
2024
DurIAN-E 2: Duration Informed Attention Network with Adaptive Variational Autoencoder and Adversarial Learning for Expressive Text-to-Speech Synthesis
2024
Top co-authors
Yuchen Hu
· 6
Eng Siong Chng
· 5
Chen Chen
· 4
Ruizhe Li
· 4
Lirong Dai
· 3
Jie Zhang
· 2
Jie Zhang
· 2
Yu Gu
· 2
Binxing Jiao
· 1
Chao Weng
· 1
Chao Weng
· 1
Chen Chen
· 1
Topics
Speech Enhancement
Speech Recognition
Multimodal Audio
Speech Translation
Text-to-Speech
Audio Generation
Audio Understanding