Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiao Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Xiao Chen
51
papers ·
908
citations ·
17
h-index
Rutgers, The State University of New Jersey · Kunming University of Science and Technology · Beijing Normal University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Conv-Transformer Transducer: Low Latency, Low Frame Rate, Streamable End-to-End Speech Recognition
2020 · 50 citations
An Investigation of Few-Shot Learning in Spoken Term Classification
2018 · 9 citations
Statistical Parametric Speech Synthesis Using Generative Adversarial Networks Under A Multi-task Learning Framework
2017 · 6 citations
VQMIVC: Vector Quantization and Mutual Information-Based Unsupervised Speech Representation Disentanglement for One-shot Voice Conversion
2021 · 4 citations
The HUAWEI Speaker Diarisation System for the VoxCeleb Speaker Diarisation Challenge
2020 · 1 citations
EditSpeech: A Text Based Speech Editing System Using Partial Inference and Bidirectional Fusion
2021 · 1 citations
Enhancing Code-switched Text-to-Speech Synthesis Capability in Large Language Models with only Monolingual Corpora
2024 · 1 citations
SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing
2026
Speech-Omni-Lite: Portable Speech Interfaces for Vision-Language Models
2026
DSA-Tokenizer: Disentangled Semantic-Acoustic Tokenization via Flow Matching-based Hierarchical Fusion
2026
PROST-LLM: Progressively Enhancing the Speech-to-Speech Translation Capability in LLMs
2026
EMOVA: Empowering Language Models to See, Hear and Speak with Vivid Emotions
2024
Dual-label Deep LSTM Dereverberation For Speaker Verification
2018
A Streaming End-to-End Framework For Spoken Language Understanding
2021
Improving End-to-End Speech Processing by Efficient Text Data Utilization with Latent Synthesis
2023
Top co-authors
Daxin Tan
· 9
Dehua Tao
· 4
Jing Xu
· 3
Dingdong Wang
· 2
Hanlin Zhang
· 2
Haochen Tan
· 2
Jiaqi Wang
· 2
Kai Chen
· 2
Lanqing Hong
· 2
Linqi Song
· 2
Liqun Deng
· 2
Wenyong Huang
· 2
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Speech Enhancement
Audio Understanding
eess.AS
Speaker Analysis
cs.SD
Speech Translation
cs.AI