Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jianwei Yu — most-cited papers & profile · Speech Audio
← authors
·
overview
Jianwei Yu
16
papers ·
80
citations ·
0
h-index
Chinese University of Hong Kong, Shenzhen
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MoonCast: High-Quality Zero-Shot Podcast Generation
2025 · 24 citations
VibeVoice Technical Report
2025 · 23 citations
Mixed Precision Low-bit Quantization of Neural Network Language Models for Speech Recognition
2021 · 17 citations
SongBloom: Coherent Song Generation via Interleaved Autoregressive Sketching and Diffusion Refinement
2025 · 16 citations
CoVoMix2: Advancing Zero-Shot Dialogue Generation with Fully Non-Autoregressive Flow Matching
2025
Pseudo-Autoregressive Neural Codec Language Models for Efficient Zero-Shot Text-to-Speech Synthesis
2025
Kimi-Audio Technical Report
2025
Improving Mandarin End-to-End Speech Recognition with Word N-gram Language Model
2022
Spectro-Temporal Deep Features for Disordered Speech Assessment and Recognition
2022
Recent Progress in the CUHK Dysarthric Speech Recognition System
2022
Integrating Lattice-Free MMI into End-to-End Speech Recognition
2022
Interleaved Speech-Text Language Models for Simple Streaming Text-to-Speech Synthesis
2024
Top co-authors
Jinyu Li
· 3
Shoukang Hu
· 3
Chao Weng
· 2
Dongchao Yang
· 2
Helen Meng
· 2
Jinchuan Tian
· 2
Lingwei Meng
· 2
Mengzhe Geng
· 2
Shansong Liu
· 2
Shujie Liu
· 2
Xinyu Zhou
· 2
Xunying Liu
· 2
Topics
Speech Recognition
Audio Generation
Text-to-Speech
Speech Translation
Music Generation
Multimodal Audio
Voice Cloning
Audio Understanding
Speech Enhancement
Speaker Analysis