Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jinglin Liu — most-cited papers & profile · Speech Audio
← authors
·
overview
Jinglin Liu
16
papers ·
240
citations ·
29
h-index
Suzhou University of Science and Technology · Suzhou Research Institute
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Multi-Singer: Fast Multi-Singer Singing Voice Vocoder With A Large-Scale Corpus
2021 · 72 citations
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism
2021 · 35 citations
GenerSpeech: Towards Style Transfer for Generalizable Out-Of-Domain Text-to-Speech
2022 · 21 citations
ProDiff: Progressive Fast Diffusion Model For High-Quality Text-to-Speech
2022 · 21 citations
RMSSinger: Realistic-Music-Score based Singing Voice Synthesis
2023 · 18 citations
TranSpeech: Speech-to-Speech Translation With Bilateral Perturbation
2022 · 17 citations
Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
2023 · 16 citations
Learning the Beauty in Songs: Neural Singing Voice Beautifier
2022 · 13 citations
Mega-TTS 2: Boosting Prompting Mechanisms for Zero-Shot Speech Synthesis
2023 · 8 citations
MR-SVS: Singing Voice Synthesis with Multi-Reference Encoder
2022 · 7 citations
DenoiSpeech: Denoising Text to Speech with Frame-Level Noise Modeling
2020 · 6 citations
PortaSpeech: Portable and High-Quality Generative Text-to-Speech
2021 · 6 citations
EMOVIE: A Mandarin Emotion Speech Dataset with a Simple Emotional Text-to-Speech Model
2021
AlignSTS: Speech-to-Singing Conversion via Cross-Modal Alignment
2023
AV-TranSpeech: Audio-Visual Robust Speech-to-Speech Translation
2023
Top co-authors
Zhou Zhao
· 15
Rongjie Huang
· 9
Yi Ren
· 9
Chenye Cui
· 5
Jinzheng He
· 5
Yi Ren
· 5
Zhenhui Ye
· 5
Huadai Liu
· 4
Xiang Yin
· 4
Chen Zhang
· 3
Lichao Zhang
· 3
Ziyue Jiang
· 3
Topics
Audio Generation
Text-to-Speech
Music Generation
Voice Cloning
Speech Recognition
Speech Enhancement
Speech Translation
Audio Understanding
Multimodal Audio