Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Nanxin Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Nanxin Chen
15
papers ·
247
citations ·
6
h-index
Google (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
ESPnet: End-to-End Speech Processing Toolkit
2018 · 74 citations
Listen and Fill in the Missing Letters: Non-Autoregressive Transformer for Speech Recognition
2019 · 53 citations
WaveGrad: Estimating Gradients for Waveform Generation
2020 · 44 citations
SLM: Bridge the thin gap between speech and text foundation models
2023 · 28 citations
Zero-Shot Multi-Speaker Text-To-Speech with State-of-the-art Neural Speaker Embeddings
2019 · 9 citations
Residual Adapters for Few-Shot Text-to-Speech Speaker Adaptation
2022 · 9 citations
Robust Training of Vector Quantized Bottleneck Models
2020 · 8 citations
A Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation
2021 · 7 citations
Mask CTC: Non-Autoregressive End-to-End ASR with CTC and Mask Predict
2020 · 6 citations
WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis
2021 · 5 citations
Efficient Adapters for Giant Speech Models
2023 · 4 citations
Focus on the present: a regularization method for the ASR source-target attention layer
2020
Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR
2022
From English to More Languages: Parameter-Efficient Model Reprogramming for Cross-Lingual Speech Recognition
2023
How to Estimate Model Transferability of Pre-Trained Speech Models?
2023
Top co-authors
Shinji Watanabe
· 4
Yu Zhang
· 4
Heiga Zen
· 3
Najim Dehak
· 3
Bo Li
· 2
Chao-Han Huck Yang
· 2
Chung-Cheng Chiu
· 2
Hagen Soltau
· 2
Izhak Shafran
· 2
Jes\'us Villalba
· 2
Mohammad Norouzi
· 2
Rohit Prabhavalkar
· 2
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Generation
Audio Understanding
Voice Cloning
Speaker Analysis
Music Generation
Speech Enhancement
Multimodal Audio