Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ziyue Jiang — most-cited papers & profile · Speech Audio
← authors
·
overview
Ziyue Jiang
21
papers ·
64
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
2023 · 16 citations
TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis
2025 · 12 citations
FluentSpeech: Stutter-Oriented Automatic Speech Editing with Context-Aware Diffusion Models
2023 · 9 citations
Mega-TTS 2: Boosting Prompting Mechanisms for Zero-Shot Speech Synthesis
2023 · 8 citations
GTSinger: A Global Multi-Technique Singing Corpus with Realistic Music Scores for All Singing Tasks
2024 · 5 citations
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
2025 · 4 citations
TCSinger: Zero-Shot Singing Voice Synthesis with Style Transfer and Multi-Level Style Control
2024 · 4 citations
Versatile Framework for Song Generation with Prompt-based Control
2025 · 2 citations
Make-A-Voice: Unified Voice Synthesis With Discrete Representation
2023 · 2 citations
MobileSpeech: A Fast and High-Fidelity Framework for Mobile Zero-Shot Text-to-Speech
2024 · 1 citations
Language-Codec: Bridging Discrete Codec Representations and Speech Language Models
2024 · 1 citations
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models
2025
DiTReducio: A Training-Free Acceleration for DiT-Based TTS via Progressive Calibration
2025
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching
2025
Discl-VC: Disentangled Discrete Tokens and In-Context Learning for Controllable Zero-Shot Voice Conversion
2025
Top co-authors
Zhou Zhao
· 14
Shengpeng Ji
· 8
Jialong Zuo
· 7
Zhenhui Ye
· 6
Qian Yang
· 4
Ruiqi Li
· 4
Xiang Yin
· 4
Xize Cheng
· 4
Zhou Zhao
· 4
Changhao Pan
· 3
Chen Zhang
· 3
Jinglin Liu
· 3
Topics
Audio Generation
Voice Cloning
Text-to-Speech
Speech Recognition
Music Generation
Audio Understanding
Multimodal Audio
Speech Enhancement
Speaker Analysis
Speech Translation