Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jingbei Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Jingbei Li
12
papers ·
17
citations ·
9
h-index
University of Hong Kong · Tsinghua University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Emotion controllable speech synthesis using emotion-unlabeled dataset with the assistance of cross-domain speech emotion recognition
2020 · 4 citations
Towards Multi-Scale Style Control for Expressive Speech Synthesis
2021 · 4 citations
Adversarially learning disentangled speech representations for robust multi-factor voice conversion
2021 · 3 citations
DiffCSS: Diverse and Expressive Conversational Speech Synthesis with Diffusion Models
2025 · 2 citations
Syntactic representation learning for neural network based TTS with syntactic parse tree traversal
2020 · 1 citations
Enhancing Speaking Styles in Conversational Text-to-Speech Synthesis with Graph-based Multi-modal Context Modeling
2021 · 1 citations
NeuFA: Neural Network Based End-to-End Forced Alignment with Bidirectional Attention Mechanism
2022 · 1 citations
DiCLET-TTS: Diffusion Model based Cross-lingual Emotion Transfer for Text-to-Speech -- A Study between English and Mandarin
2023 · 1 citations
Step-Audio 2 Technical Report
2025
Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model
2025
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
2025
Enhancing Word-Level Semantic Representation via Dependency Structure for Expressive Text-to-Speech Synthesis
2021
Top co-authors
Zhiyong Wu
· 8
Helen Meng
· 7
Changhe Song
· 4
Bingxin Li
· 3
Bin Wang
· 3
Binxing Jiao
· 3
Bo Li
· 3
Boyong Wu
· 3
Buyun Ma
· 3
Changxin Miao
· 3
Changyi Wan
· 3
Chao Yan
· 3
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Multimodal Audio
Audio Understanding
Speech Translation
Music Generation
Voice Cloning
Speech Enhancement
Speaker Analysis