Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jian Cong — most-cited papers & profile · Speech Audio
← authors
·
overview
Jian Cong
15
papers ·
86
citations ·
10
h-index
Mongolian University of Science and Technology · Inner Mongolia University of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
NaturalSpeech: End-to-End Text to Speech Synthesis with Human-Level Quality
2022 · 35 citations
Data Efficient Voice Cloning from Noisy Samples with Domain Adversarial Training
2020 · 32 citations
VISinger: Variational Inference with Adversarial Learning for End-to-End Singing Voice Synthesis
2021 · 10 citations
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
2024 · 5 citations
Glow-WaveGAN: Learning Speech Representations from GAN-based Variational Auto-Encoder For High Fidelity Flow-based Speech Synthesis
2021 · 3 citations
DiCLET-TTS: Diffusion Model based Cross-lingual Emotion Transfer for Text-to-Speech -- A Study between English and Mandarin
2023 · 1 citations
MagiCodec: Simple Masked Gaussian-Injected Codec for High-Fidelity Reconstruction and Generation
2025
DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation
2025
Controllable Context-aware Conversational Speech Synthesis
2021
AdaVITS: Tiny VITS for Low Computing Resource Speaker Adaptation
2022
Glow-WaveGAN 2: High-quality Zero-shot Text-to-speech Synthesis and Any-to-any Voice Conversion
2022
Robust MelGAN: A robust universal neural vocoder for high-fidelity TTS
2022
DSPGAN: a GAN-based universal vocoder for high-fidelity TTS by time-frequency domain supervision from DSP
2022
U-Style: Cascading U-nets with Multi-level Speaker and Style Modeling for Zero-Shot Voice Cloning
2023
Language Model Can Listen While Speaking
2024
Top co-authors
Lei Xie
· 10
Yuping Wang
· 6
Dan Su
· 4
Jiawei Chen
· 4
Shan Yang
· 4
Yongmao Zhang
· 4
Yuxuan Wang
· 4
Zhuo Chen
· 4
Chenpeng Du
· 3
Dongya Jia
· 3
Jian Wu
· 3
Kun Song
· 3
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Speech Enhancement
Music Generation
Voice Cloning
Speech Translation
Speaker Analysis
Audio Understanding
Multimodal Audio