Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Heiga Zen — most-cited papers & profile · Speech Audio
← authors
·
overview
Heiga Zen
30
papers ·
4617
citations ·
45
h-index
Google (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
WaveNet: A Generative Model for Raw Audio
2016 · 3624 citations
Parallel WaveNet: Fast High-Fidelity Speech Synthesis
2017 · 344 citations
Fully-hierarchical fine-grained prosody modeling for interpretable speech synthesis
2020 · 93 citations
Sample Efficient Adaptive Text-to-Speech
2018 · 75 citations
Non-Attentive Tacotron: Robust and Controllable Neural TTS Synthesis Including Unsupervised Duration Modeling
2020 · 73 citations
MAESTRO: Matched Speech Text Representations through Modality Matching
2022 · 71 citations
PnG BERT: Augmented BERT on Phonemes and Graphemes for Neural TTS
2021 · 61 citations
Hierarchical Generative Modeling for Controllable Speech Synthesis
2018 · 45 citations
WaveGrad: Estimating Gradients for Waveform Generation
2020 · 44 citations
Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning
2019 · 26 citations
LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
2019 · 23 citations
Parallel Tacotron: Non-Autoregressive and Controllable TTS
2020 · 21 citations
Generating diverse and natural text-to-speech samples using a quantized fine-grained VAE and auto-regressive prosody prior
2020 · 15 citations
Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling
2021 · 11 citations
Training Text-To-Speech Systems From Synthetic Data: A Practical Approach For Accent Transfer Tasks
2022 · 11 citations
Top co-authors
Ye Jia
· 9
Yonghui Wu
· 9
Ron J. Weiss
· 7
Jonathan Shen
· 6
Michiel Bacchiani
· 6
Yuma Koizumi
· 6
Yu Zhang
· 6
Yu Zhang
· 6
Bhuvana Ramabhadran
· 5
Nobuyuki Morioka
· 5
Ankur Bapna
· 4
Isaac Elias
· 4
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Speech Enhancement
Speech Translation
Music Generation
Voice Cloning
Speaker Analysis
Audio Understanding
Multimodal Audio