Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Rongjie Huang — most-cited papers & profile · AI for Code
← authors
·
overview
Rongjie Huang
36
papers ·
265
citations ·
17
h-index
Zhengzhou University of Light Industry · First Affiliated Hospital of GuangXi Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Multi-Singer: Fast Multi-Singer Singing Voice Vocoder With A Large-Scale Corpus
2021 · 72 citations
FastDiff: A Fast Conditional Diffusion Model for High-Quality Speech Synthesis
2022 · 28 citations
GenerSpeech: Towards Style Transfer for Generalizable Out-Of-Domain Text-to-Speech
2022 · 21 citations
ProDiff: Progressive Fast Diffusion Model For High-Quality Text-to-Speech
2022 · 21 citations
HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
2023 · 19 citations
RMSSinger: Realistic-Music-Score based Singing Voice Synthesis
2023 · 18 citations
TranSpeech: Speech-to-Speech Translation With Bilateral Perturbation
2022 · 17 citations
Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
2023 · 16 citations
UniAudio: An Audio Foundation Model Toward Universal Audio Generation
2023 · 14 citations
FluentSpeech: Stutter-Oriented Automatic Speech Editing with Context-Aware Diffusion Models
2023 · 9 citations
TechSinger: Technique Controllable Multilingual Singing Voice Synthesis via Flow Matching
2025 · 7 citations
InstructTTS: Modelling Expressive TTS in Discrete Latent Space with Natural Language Style Prompt
2023 · 5 citations
Prompt-Singer: Controllable Singing-Voice-Synthesis with Natural Language Prompt
2024 · 4 citations
TCSinger: Zero-Shot Singing Voice Synthesis with Style Transfer and Multi-Level Style Control
2024 · 4 citations
Versatile Framework for Song Generation with Prompt-based Control
2025 · 2 citations
Topics
Audio Generation
Text-to-Speech
Music Generation
Multimodal Audio
Speech Recognition
Voice Cloning
Audio Understanding
Speech Translation
Speech Enhancement
Model Architecture