Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chen Xu — most-cited papers & profile · Speech Audio
← authors
·
overview
Chen Xu
68
papers ·
177
citations ·
8
h-index
Minzu University of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Stacked Acoustic-and-Textual Encoding: Integrating the Pre-trained Models into Speech Translation Encoders
2021 · 6 citations
The NiuTrans End-to-End Speech Translation System for IWSLT 2021 Offline Task
2021 · 5 citations
PolyVoice: Language Models for Speech to Speech Translation
2023 · 5 citations
Rethinking and Improving Multi-task Learning for End-to-end Speech Translation
2023 · 4 citations
Enhancing Speech Large Language Models with Prompt-Aware Mixture of Audio Encoders
2025 · 3 citations
Improving End-to-end Speech Translation by Leveraging Auxiliary Speech and Text Data
2022 · 3 citations
Recent Advances in Direct Speech-to-text Translation
2023 · 2 citations
Soft Alignment of Modality Space for End-to-end Speech Translation
2023 · 1 citations
WaveEx: Accelerating Flow Matching-based Speech Generation via Wavelet-guided Extrapolation
2026
Step-Audio 2 Technical Report
2025
Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model
2025
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
2025
Bridging the Granularity Gap for Acoustic Modeling
2023
CTC-based Non-autoregressive Speech Translation
2023
Bridging the Gaps of Both Modality and Language: Synchronous Bilingual CTC for Speech Translation and Speech Recognition
2023
Top co-authors
Jingbo Zhu
· 14
Tong Xiao
· 14
Xiaoqian Liu
· 8
Yuhao Zhang
· 8
Chunliang Zhang
· 4
Bingxin Li
· 3
Bin Wang
· 3
Binxing Jiao
· 3
Bo Li
· 3
Changxin Miao
· 3
Changyi Wan
· 3
Chao Yan
· 3
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Generation
Multimodal Audio
Audio Understanding
Music Generation
Speaker Analysis
Speech Enhancement
cs.CL