Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jingbin Hu — most-cited papers & profile · Speech Audio
← authors
·
overview
Jingbin Hu
8
papers ·
2
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
HiStyle: Hierarchical Style Embedding Predictor for Text-Prompt-Guided Controllable Speech Synthesis
2025 · 2 citations
Seeing the Context: Rich Visual Context-Aware Speech Recognition via Multimodal Reasoning
2026
OmniCodec: Low Frame Rate Universal Audio Codec with Semantic-Acoustic Disentanglement
2026
VoiceSculptor: Your Voice, Designed By You
2026
WenetSpeech-Wu: Datasets, Benchmarks, and Models for a Unified Chinese Wu Dialect Speech Processing Ecosystem
2026
XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation
2025
FleSpeech: Flexibly Controllable Speech Generation with Various Prompts
2025
Top co-authors
Chengyou Wang
· 2
Lei Xie
· 2
Pengyuan Xie
· 2
Wenhao Li
· 2
Bengu Wu
· 1
Binbin Zhang
· 1
Binbin Zhang
· 1
Bingshen Mu
· 1
Bingshen Mu
· 1
Chengyou Wang
· 1
Chengyou Wang
· 1
Chuang Ding
· 1
Topics
Text-to-Speech
Audio Generation
Speech Translation
Speech Recognition
Audio Understanding
Music Generation
Multimodal Audio
Voice Cloning
Speaker Analysis