Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Wei Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Wei Chen
182
papers ·
539
citations ·
16
h-index
Roma Tre University · Microsoft Research Asia (China) · Guilin University of Electronic Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improving Generalization of Transformer for Speech Recognition with Parallel Schedule Sampling and Relative Positional Embedding
2019 · 23 citations
PriorGrad: Improving Conditional Denoising Diffusion Models with Data-Dependent Adaptive Prior
2021 · 23 citations
Multi-band MelGAN: Faster Waveform Generation for High-Quality Text-to-Speech
2020 · 21 citations
Improving the Robustness of Speech Translation
2018 · 14 citations
WNARS: WFST based Non-autoregressive Streaming End-to-End Speech Recognition
2021 · 9 citations
Exploring RNN-Transducer for Chinese Speech Recognition
2018 · 6 citations
An Online Attention-based Model for Speech Recognition
2018 · 6 citations
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024 · 5 citations
Extending Recurrent Neural Aligner for Streaming End-to-End Speech Recognition in Mandarin
2018 · 2 citations
AutoStyle-TTS: Retrieval-Augmented Generation based Automatic Style Matching Text-to-Speech Synthesis
2025 · 1 citations
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
2025 · 1 citations
Modality Attention for End-to-End Audio-visual Speech Recognition
2018 · 1 citations
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
2024 · 1 citations
Multi-Loss Learning for Speech Emotion Recognition with Energy-Adaptive Mixup and Frame-Level Attention
2025
DAFMSVC: One-Shot Singing Voice Conversion with Dual Attention Mechanism and Flow Matching
2025
Top co-authors
Pan Zhou
· 7
Lei Xie
· 5
Zhiyong Wu
· 4
Dan Luo
· 2
Fan Fan
· 2
He Wang
· 2
Jing Yang
· 2
Pengcheng Guo
· 2
Xiang Li
· 2
Xin Xu
· 2
Zhuo Chen
· 2
Zhuo Wang
· 2
Topics
Speech Recognition
Speech Translation
Audio Understanding
Text-to-Speech
Music Generation
Audio Generation
Speech Enhancement
Voice Cloning
cs.SD
cs.AI