Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuancheng Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Yuancheng Wang
16
papers ·
39
citations ·
26
h-index
Zhongda Hospital Southeast University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
2024 · 20 citations
AnyEnhance: A Unified Generative Model with Prompt-Guidance and Self-Critic for Voice Enhancement
2025 · 4 citations
Amphion: An Open-Source Audio, Music and Speech Generation Toolkit
2023 · 3 citations
Calibration of a two-state pitch-wise HMM method for note segmentation in Automatic Music Transcription systems
2017 · 2 citations
RALL-E: Robust Codec Language Modeling with Chain-of-Thought Prompting for Text-to-Speech Synthesis
2024 · 2 citations
Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation
2024 · 2 citations
Emilia: A Large-Scale, Extensive, Multilingual, and Diverse Dataset for Speech Generation
2025 · 1 citations
SpeechJudge: Towards Human-Level Judgment for Speech Naturalness
2025
Vevo2: A Unified and Controllable Framework for Speech and Singing Voice Generation
2025
Advancing Zero-shot Text-to-Speech Intelligibility across Diverse Domains via Preference Alignment
2025
Metis: A Foundation Speech Generation Model with Masked Generative Pre-training
2025
Overview of the Amphion Toolkit (v0.2)
2025
MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
2024
Debatts: Zero-Shot Debating Text-to-Speech Synthesis
2024
Noro: Noise-Robust One-shot Voice Conversion with Hidden Speaker Representation Learning
2024
Top co-authors
Zhizheng Wu
· 10
Chaoren Wang
· 7
Haorui He
· 5
Junan Zhang
· 3
Yicheng Gu
· 3
Chen Yang
· 2
Detai Xin
· 2
Dongchao Yang
· 2
Dongya Jia
· 2
Haotian Guo
· 2
Hua Hua
· 2
Jiachen Zheng
· 2
Topics
Audio Generation
Text-to-Speech
Music Generation
Speech Translation
Speech Enhancement
Speech Recognition
Voice Cloning
cs.IR
cs.SC
Speaker Analysis