Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ning Cheng — most-cited papers & profile · Multimodal
← authors
·
overview
Ning Cheng
52
papers ·
256
citations ·
18
h-index
Chinese Academy of Sciences · Institute of High Energy Physics
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
AlignTTS: Efficient Feed-Forward Text-to-Speech System without Explicit Alignment
2020 · 64 citations
Applying Wav2vec2.0 to Speech Recognition in Various Low-resource Languages
2020 · 57 citations
GraphTTS: graph-to-sequence modelling in neural text-to-speech
2020 · 14 citations
TDASS: Target Domain Adaptation Speech Synthesis Framework for Multi-speaker Low-Resource TTS
2022 · 13 citations
SUSing: SU-net for Singing Voice Synthesis
2022 · 13 citations
Prosody Learning Mechanism for Speech Synthesis System Without Text Length Limit
2020 · 10 citations
Voice Conversion with Denoising Diffusion Probabilistic GAN Models
2023 · 9 citations
Voice Conversion with Denoising Diffusion Probabilistic GAN Models
2023 · 9 citations
EmoTalker: Emotionally Editable Talking Face Generation via Diffusion Model
2024 · 9 citations
Unidirectional Memory-Self-Attention Transducer for Online Speech Recognition
2021 · 7 citations
Large-scale Transfer Learning for Low-resource Spoken Language Understanding
2020 · 6 citations
Semi-Supervised Learning Based on Reference Model for Low-resource TTS
2022 · 6 citations
Dropout Regularization for Self-Supervised Learning of Transformer Encoder Speech Representation
2021 · 4 citations
EAD-VC: Enhancing Speech Auto-Disentanglement for Voice Conversion with IFUB Estimator and Joint Text-Guided Consistent Learning
2024 · 3 citations
Touch100k: A Large-Scale Touch-Language-Vision Dataset for Touch-Centric Multimodal Representation
2024 · 3 citations
Topics
Audio Generation
Speech Recognition
Text-to-Speech
Speech Enhancement
Voice Cloning
Speech Translation
Speaker Analysis
Audio Understanding
Diffusion Models
Conditioning & Control