Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Bin Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Bin Zhang
56
papers ·
193
citations ·
10
h-index
Zhongyuan University of Technology · Shanghai Jiao Tong University · Zhengzhou University of Light Industry · Zhengzhou University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Phase-aware music super-resolution using generative adversarial networks
2020 · 22 citations
Qwen2.5-Omni Technical Report
2025 · 8 citations
IntrinsicVoice: Empowering LLMs with Intrinsic Real-time Voice Interaction Abilities
2024 · 1 citations
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing
2026
Qwen3-TTS Technical Report
2026
VITS-Based Singing Voice Conversion Leveraging Whisper and multi-scale F0 Modeling
2023
Top co-authors
Hangrui Hu
· 2
Jin Xu
· 2
Junyang Lin
· 2
Ting He
· 2
Xiong Wang
· 2
Zhifang Guo
· 2
Baosong Yang
· 1
Dake Guo
· 1
Dong Zhang
· 1
Hongkun Hao
· 1
Jiacheng Xu
· 1
Jialin Wang
· 1
Topics
Audio Generation
Speech Recognition
Multimodal Audio
Text-to-Speech
cs.CL
Music Generation
cs.AI
cs.SD
eess.AS
Voice Cloning