Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xie Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Xie Chen
31
papers ·
111
citations ·
12
h-index
Chinese University of Hong Kong · Shanghai Jiao Tong University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition
2021 · 43 citations
VQTTS: High-Fidelity Text-to-Speech Synthesis with Self-Supervised VQ Acoustic Feature
2022 · 43 citations
Developing Real-time Streaming Transformer Transducer for Speech Recognition on Large-scale Dataset
2020 · 9 citations
Improving Code-Switching and Named Entity Recognition in ASR with Speech Editing based Data Augmentation
2023 · 4 citations
Unlocking Temporal Flexibility: Neural Speech Codec with Variable Frame Rate
2025 · 3 citations
VALL-T: Decoder-Only Generative Transducer for Robust and Decoding-Controllable Text-to-Speech
2024 · 3 citations
ELLA-V: Stable Neural Codec Language Modeling with Alignment-guided Sequence Reordering
2024 · 2 citations
Factorized Neural Transducer for Efficient Language Model Adaptation
2021 · 1 citations
Internal Language Model Adaptation with Text-Only Data for End-to-End Speech Recognition
2021 · 1 citations
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition
2023 · 1 citations
Expressive TTS Driven by Natural Language Prompts Using Few Human Annotations
2023 · 1 citations
Accelerating Diffusion-based Text-to-Speech Model Training with Dual Modality Alignment
2025
Pseudo-Autoregressive Neural Codec Language Models for Efficient Zero-Shot Text-to-Speech Synthesis
2025
SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation
2025
An Adapter based Multi-label Pre-training for Speech Separation and Enhancement
2022
Top co-authors
Yiwei Guo
· 11
Chenpeng Du
· 9
Ziyang Ma
· 9
Kai Yu
· 7
Jinyu Li
· 6
Kai Yu
· 5
Zhisheng Zheng
· 4
Changli Tang
· 3
Hankun Wang
· 3
Shujie Liu
· 3
Yujin Wang
· 3
Zhong Meng
· 3
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Speech Translation
Audio Understanding
Speech Enhancement
Speaker Analysis
Voice Cloning
Multimodal Audio
Music Generation