Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yang Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Yang Zhang
242
papers ·
8045
citations ·
46
h-index
Nanjing University of Aeronautics and Astronautics
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
AUTOVC: Zero-Shot Voice Style Transfer with Only Autoencoder Loss
2019 · 195 citations
Stochastic Gradient Methods with Layer-wise Adaptive Moments for Training of Deep Networks
2019 · 88 citations
Hi-Fi Multi-Speaker English TTS Dataset
2021 · 69 citations
Unsupervised Speech Decomposition via Triple Information Bottleneck
2020 · 43 citations
QuartzNet: Deep Automatic Speech Recognition with 1D Time-Channel Separable Convolutions
2019 · 31 citations
ContentVec: An Improved Self-Supervised Speech Representation by Disentangling Speakers
2022 · 24 citations
Deep LSTM for Large Vocabulary Continuous Speech Recognition
2017 · 23 citations
ByteSing: A Chinese Singing Voice Synthesis System Using Duration Allocated Encoder-Decoder Acoustic Models and WaveRNN Vocoders
2020 · 17 citations
Conformer-based Target-Speaker Automatic Speech Recognition for Single-Channel Audio
2023 · 16 citations
Unified Mandarin TTS Front-end Based on Distilled BERT Model
2020 · 15 citations
PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition
2021 · 12 citations
Global Rhythm Style Transfer Without Text Transcriptions
2021 · 12 citations
Shallow Fusion of Weighted Finite-State Transducer and Language Model for Text Normalization
2022 · 11 citations
Speech Denoising with Auditory Models
2020 · 8 citations
A Unified Transformer-based Framework for Duplex Text Normalization
2021 · 7 citations
Top co-authors
Kaizhi Qian
· 14
Shiyu Chang
· 9
Mark Hasegawa-Johnson
· 8
Boris Ginsburg
· 7
David Cox
· 5
Yuxuan Wang
· 5
Zejun Ma
· 5
Evelina Bakhturina
· 4
Jun Zhang
· 4
Vitaly Lavrukhin
· 4
Yi He
· 3
Yuping Wang
· 3
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Understanding
Speaker Analysis
Audio Generation
Speech Enhancement
Voice Cloning
Music Generation
Multimodal Audio