Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Dongya Jia — most-cited papers & profile · Speech Audio
← authors
·
overview
Dongya Jia
11
papers ·
7
citations ·
26
h-index
ZTE (China) · National Cancer Institute · Center for Cancer Research
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
2024 · 5 citations
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
2024 · 2 citations
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling
2026
On the Distillation Loss Functions of Speech VAE for Unified Reconstruction, Understanding, and Generation
2026
SpeechJudge: Towards Human-Level Judgment for Speech Naturalness
2025
Vevo2: A Unified and Controllable Framework for Speech and Singing Voice Generation
2025
MagiCodec: Simple Masked Gaussian-Injected Codec for High-Fidelity Reconstruction and Generation
2025
DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation
2025
Zero-Shot Accent Conversion using Pseudo Siamese Disentanglement Network
2022
VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
2024
Top co-authors
Yuping Wang
· 6
Yuanzhe Chen
· 4
Chumin Li
· 3
Jian Cong
· 3
Jiawei Chen
· 3
Yuxuan Wang
· 3
Zhuo Chen
· 3
Chaoren Wang
· 2
Chao Yao
· 2
Chenpeng Du
· 2
Dejian Zhong
· 2
Hui Li
· 2
Topics
Audio Generation
Text-to-Speech
Music Generation
Speech Recognition
Speech Translation
Speech Enhancement
Voice Cloning
Audio Understanding
Multimodal Audio