Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Rongjie Huang โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Rongjie Huang
27
papers ยท
254
citations ยท
17
h-index
Zhengzhou University of Light Industry ยท First Affiliated Hospital of GuangXi Medical University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Multi-Singer: Fast Multi-Singer Singing Voice Vocoder With A Large-Scale Corpus
2021 ยท 72 citations
FastDiff: A Fast Conditional Diffusion Model for High-Quality Speech Synthesis
2022 ยท 28 citations
GenerSpeech: Towards Style Transfer for Generalizable Out-Of-Domain Text-to-Speech
2022 ยท 21 citations
ProDiff: Progressive Fast Diffusion Model For High-Quality Text-to-Speech
2022 ยท 21 citations
HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
2023 ยท 19 citations
RMSSinger: Realistic-Music-Score based Singing Voice Synthesis
2023 ยท 18 citations
TranSpeech: Speech-to-Speech Translation With Bilateral Perturbation
2022 ยท 17 citations
Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
2023 ยท 16 citations
UniAudio: An Audio Foundation Model Toward Universal Audio Generation
2023 ยท 14 citations
FluentSpeech: Stutter-Oriented Automatic Speech Editing with Context-Aware Diffusion Models
2023 ยท 9 citations
InstructTTS: Modelling Expressive TTS in Discrete Latent Space with Natural Language Style Prompt
2023 ยท 5 citations
Prompt-Singer: Controllable Singing-Voice-Synthesis with Natural Language Prompt
2024 ยท 4 citations
Make-A-Voice: Unified Voice Synthesis With Discrete Representation
2023 ยท 2 citations
Wav2SQL: Direct Generalizable Speech-To-SQL Parsing
2023 ยท 1 citations
Speech-to-Speech Translation with Discrete-Unit-Based Style Transfer
2023 ยท 1 citations
Top co-authors
Zhou Zhao
ยท 21
Jinglin Liu
ยท 9
Yi Ren
ยท 7
Huadai Liu
ยท 6
Chenye Cui
ยท 5
Jinzheng He
ยท 5
Yongqi Wang
ยท 5
Zhenhui Ye
ยท 5
Dongchao Yang
ยท 4
Xize Cheng
ยท 4
Zhiqing Hong
ยท 4
Chao Weng
ยท 3
Topics
Audio Generation
Text-to-Speech
Music Generation
Speech Recognition
Multimodal Audio
Speech Translation
Voice Cloning
Audio Understanding
Speech Enhancement