Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhifeng Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Zhifeng Chen
30
papers ·
1727
citations ·
37
h-index
Zhejiang Sci-Tech University · Zhejiang Lab
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis
2018 · 435 citations
A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
2020 · 202 citations
Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
2017 · 184 citations
Large-Scale Multilingual Speech Recognition with a Streaming End-to-End Model
2019 · 158 citations
Tacotron: Towards End-to-End Speech Synthesis
2017 · 152 citations
A comparison of end-to-end models for long-form speech recognition
2019 · 85 citations
Sequence-to-Sequence Models Can Directly Translate Foreign Speech
2017 · 55 citations
Noise2Music: Text-conditioned Music Generation with Diffusion Models
2023 · 50 citations
Hierarchical Generative Modeling for Controllable Speech Synthesis
2018 · 45 citations
Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning
2019 · 26 citations
LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
2019 · 23 citations
Direct speech-to-speech translation with a sequence-to-sequence model
2019 · 22 citations
Massively Multilingual Shallow Fusion with Large Language Models
2023 · 14 citations
Multi-Dialect Speech Recognition With A Single Sequence-To-Sequence Model
2017 · 10 citations
State-of-the-art Speech Recognition With Sequence-to-Sequence Models
2017
Top co-authors
Yonghui Wu
· 15
Yu Zhang
· 9
Bo Li
· 5
Chung-Cheng Chiu
· 5
Ruoming Pang
· 5
Navdeep Jaitly
· 4
Bhuvana Ramabhadran
· 3
Heiga Zen
· 3
Jonathan Shen
· 3
Wei Han
· 3
Yuxuan Wang
· 3
Jiahui Yu
· 2
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Generation
Voice Cloning
Speaker Analysis
Speech Enhancement
Music Generation
Multimodal Audio