Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiaoyu Yang — most-cited papers & profile · Speech Audio
← authors
·
overview
Xiaoyu Yang
21
papers ·
50
citations ·
0
h-index
Jiangnan University · Xidian University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Zipformer: A faster and better encoder for automatic speech recognition
2023 · 28 citations
SALMONN-omni: A Standalone Speech LLM without Codec Injection for Full-duplex Conversation
2025 · 16 citations
Knowledge Distillation for Neural Transducers from Large Self-Supervised Pre-trained Models
2021 · 1 citations
Delay-penalized transducer for low-latency streaming ASR
2022 · 1 citations
Knowledge Distillation from Multiple Foundation Models for End-to-End Speech Recognition
2023 · 1 citations
Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context
2023 · 1 citations
SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
2024 · 1 citations
SPEAR: A Unified SSL Framework for Learning Speech and Audio Representations
2025
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning
2024
CR-CTC: Consistency regularization on CTC for improved speech recognition
2024
MT2KD: Towards A General-Purpose Encoder for Speech, Speaker, and Audio Events
2024
Fast and parallel decoding for transducer
2022
Blank-regularized CTC for Frame Skipping in Neural Transducer
2023
PromptASR for contextualized ASR with controllable style
2023
LibriheavyMix: A 20,000-Hour Dataset for Single-Channel Reverberant Multi-Talker Speech Separation, ASR and Speaker Diarization
2024
Top co-authors
Wei Kang
· 9
Zengwei Yao
· 9
Yifan Yang
· 7
Chao Zhang
· 5
Zengrui Jin
· 5
Qiujia Li
· 3
Guangzhi Sun
· 2
Jun Zhang
· 2
Lu Lu
· 2
Philip C. Woodland
· 2
Phil Woodland
· 2
Siyin Wang
· 2
Topics
Speech Recognition
Audio Understanding
Speech Translation
Speech Enhancement
Multimodal Audio
Speaker Analysis
Text-to-Speech
eess.AS
Audio Generation