Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Xie Chen โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Xie Chen
104
papers ยท
201
citations ยท
12
h-index
Chinese University of Hong Kong ยท Shanghai Jiao Tong University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition
2021 ยท 43 citations
VQTTS: High-Fidelity Text-to-Speech Synthesis with Self-Supervised VQ Acoustic Feature
2022 ยท 43 citations
Phonetic and Graphemic Systems for Multi-Genre Broadcast Transcription
2018 ยท 16 citations
Minimum Word Error Rate Training with Language Model Fusion for End-to-End Speech Recognition
2021 ยท 14 citations
Developing Real-time Streaming Transformer Transducer for Speech Recognition on Large-scale Dataset
2020 ยท 9 citations
Long-span language modeling for speech recognition
2019 ยท 8 citations
Semantic-VAE: Semantic-Alignment Latent Representation for Better Speech Synthesis
2025 ยท 6 citations
SoulX-Podcast: Towards Realistic Long-form Podcasts with Dialectal and Paralinguistic Diversity
2025 ยท 5 citations
Enhancing Speech-to-Speech Dialogue Modeling with End-to-End Retrieval-Augmented Generation
2025 ยท 5 citations
AUV: Teaching Audio Universal Vector Quantization with Single Nested Codebook
2025 ยท 4 citations
LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition
2020 ยท 4 citations
UniCATS: A Unified Context-Aware Text-to-Speech Framework with Contextual VQ-Diffusion and Vocoding
2023 ยท 4 citations
Improving Code-Switching and Named Entity Recognition in ASR with Speech Editing based Data Augmentation
2023 ยท 4 citations
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
2024 ยท 4 citations
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy
2025 ยท 3 citations
Top co-authors
Ziyang Ma
ยท 41
Kai Yu
ยท 33
Yifan Yang
ยท 15
Guanrou Yang
ยท 12
Jinyu Li
ยท 10
Shuai Wang
ยท 7
Zhuo Chen
ยท 7
Shiliang Zhang
ยท 6
Shujie Liu
ยท 6
Bohan Li
ยท 5
Lei Xie
ยท 5
Xunying Liu
ยท 5
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Speech Translation
cs.SD
eess.AS
Audio Understanding
Speech Enhancement
cs.AI
cs.CL