Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ye Jia — most-cited papers & profile · Speech Audio
← authors
·
overview
Ye Jia
27
papers ·
1404
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis
2018 · 475 citations
Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis
2018 · 435 citations
Non-Attentive Tacotron: Robust and Controllable Neural TTS Synthesis Including Unsupervised Duration Modeling
2020 · 73 citations
PnG BERT: Augmented BERT on Phonemes and Graphemes for Neural TTS
2021 · 61 citations
mSLAM: Massively multilingual joint pre-training for speech and text
2022 · 59 citations
VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking
2018 · 50 citations
SLAM: A Unified Encoder for Speech and Language Modeling via Speech-Text Joint Pre-Training
2021 · 50 citations
Hierarchical Generative Modeling for Controllable Speech Synthesis
2018 · 45 citations
Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning
2019 · 26 citations
LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
2019 · 23 citations
Direct speech-to-speech translation with a sequence-to-sequence model
2019 · 22 citations
Parallel Tacotron: Non-Autoregressive and Controllable TTS
2020 · 21 citations
Translatotron 2: High-quality direct speech-to-speech translation with voice preservation
2021 · 21 citations
Parrotron: An End-to-End Speech-to-Speech Conversion Model and its Applications to Hearing-Impaired Speech and Speech Separation
2019 · 19 citations
Leveraging unsupervised and weakly-supervised data to improve direct speech-to-speech translation
2022 · 15 citations
Top co-authors
Yu Zhang
· 15
Yonghui Wu
· 12
Heiga Zen
· 10
Ron J. Weiss
· 8
Jonathan Shen
· 7
Melvin Johnson
· 5
Zhifeng Chen
· 5
Alexis Conneau
· 4
Ankur Bapna
· 4
Michelle Tadmor Ramanovich
· 4
Quan Wang
· 4
Colin Cherry
· 3
Topics
Text-to-Speech
Speech Recognition
Audio Generation
Speech Translation
Speech Enhancement
Voice Cloning
Multimodal Audio
Speaker Analysis
Audio Understanding