Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhiyong Wu — most-cited papers & profile · Speech Audio
← authors
·
overview
Zhiyong Wu
16
papers ·
108
citations ·
33
h-index
Shandong University of Technology · Shantou Central Hospital · Tsinghua University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving
2025 · 2 citations
E2E-VGuard: Adversarial Prevention for Production LLM-based End-To-End Speech Synthesis
2025
LSZone: A Lightweight Spatial Information Modeling Architecture for Real-time In-car Multi-zone Speech Separation
2025
LightGrad: Lightweight Diffusion Probabilistic Model for Text-to-Speech
2023
Improving Mandarin Prosodic Structure Prediction with Multi-level Contextual Information
2023
RFWave: Multi-band Rectified Flow for Audio Waveform Reconstruction
2024
VoxInstruct: Expressive Human Instruction-to-Speech Generation with Unified Multilingual Codec Language Modelling
2024
Top co-authors
Jie Chen
· 2
Xixin Wu
· 2
Binbin Zhang
· 1
Changhe Song
· 1
Chao Weng
· 1
Derui Wang
· 1
Deyi Tuo
· 1
Dongyang Dai
· 1
Fuping Pan
· 1
Helen Meng
· 1
Hui Wang
· 1
Jia Jia
· 1
Topics
Text-to-Speech
Speech Recognition
Audio Generation
Audio Understanding
Speech Translation
Music Generation
Multimodal Audio
Voice Cloning
Speech Enhancement