Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chao Weng — most-cited papers & profile · Speech Audio
← authors
·
overview
Chao Weng
36
papers ·
445
citations ·
27
h-index
Tencent (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
2021 · 201 citations
DurIAN: Duration Informed Attention Network For Multimodal Synthesis
2019 · 94 citations
Towards Robust Speaker Verification with Target Speaker Enhancement
2021 · 19 citations
HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
2023 · 19 citations
Replay and Synthetic Speech Detection with Res2net Architecture
2020 · 18 citations
Minimum Bayes Risk Training of RNN-Transducer for End-to-End Speech Recognition
2019 · 16 citations
VARA-TTS: Non-Autoregressive Text-to-Speech Synthesis based on Very Deep VAE with Residual Attention
2021 · 15 citations
DurIAN-SC: Duration Informed Attention Network based Singing Voice Conversion System
2020 · 13 citations
Learning Singing From Speech
2019 · 10 citations
DFSMN-SAN with Persistent Memory Model for Automatic Speech Recognition
2019 · 5 citations
InstructTTS: Modelling Expressive TTS in Discrete Latent Space with Natural Language Style Prompt
2023 · 5 citations
LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement
2025 · 4 citations
Non-Autoregressive Transformer ASR with CTC-Enhanced Decoder Input
2020 · 4 citations
Directional ASR: A New Paradigm for E2E Multi-Speaker Speech Recognition with Source Localization
2020 · 4 citations
Synthesising Expressiveness in Peking Opera via Duration Informed Attention Network
2019 · 3 citations
Top co-authors
Dong Yu
· 18
Dan Su
· 14
Helen Meng
· 7
Chengzhu Yu
· 6
Songxiang Liu
· 6
Heng Lu
· 5
Dongchao Yang
· 4
Jianwei Yu
· 4
Meng Yu
· 4
Shinji Watanabe
· 4
Yuexian Zou
· 4
Zhiyong Wu
· 4
Topics
Audio Generation
Speech Recognition
Text-to-Speech
Speech Enhancement
Speech Translation
Audio Understanding
Music Generation
Voice Cloning
Multimodal Audio
Speaker Analysis