Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiaofei Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Xiaofei Wang
21
papers ·
41
citations ·
47
h-index
Tianjin University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Multi-encoder multi-resolution framework for end-to-end speech recognition
2018 · 15 citations
TransVIP: Speech to Speech Translation System with Voice and Isochrony Preservation
2024 · 5 citations
Multi-Stream End-to-End Speech Recognition
2019 · 3 citations
Investigation of End-To-End Speaker-Attributed ASR for Continuous Multi-Talker Recordings
2020 · 3 citations
Large-Scale Pre-Training of End-to-End Multi-Talker ASR for Meeting Transcription with Single Distant Microphone
2021 · 3 citations
SpeechX: Neural Codec Language Model as a Versatile Speech Transformer
2023 · 3 citations
Minimum Bayes Risk Training for End-to-End Speaker-Attributed ASR
2020 · 2 citations
Exploring Methods for the Automatic Detection of Errors in Manual Transcription
2019 · 1 citations
A Comparative Study of Modular and Joint Approaches for Speaker-Attributed ASR on Monaural Long-Form Audio
2021 · 1 citations
Simulating realistic speech overlaps improves multi-talker ASR
2022 · 1 citations
DiariST: Streaming Speech Translation with Speaker Diarization
2023 · 1 citations
CoVoMix: Advancing Zero-Shot Speech Generation for Human-like Multi-talker Conversations
2024 · 1 citations
CoVoMix2: Advancing Zero-Shot Dialogue Generation with Fully Non-Autoregressive Flow Matching
2025
Towards Efficient Speech-Text Jointly Decoding within One Speech Language Model
2025
A practical two-stage training strategy for multi-stream end-to-end speech recognition
2019
Top co-authors
Jinyu Li
· 9
Naoyuki Kanda
· 9
Takuya Yoshioka
· 8
Dongmei Wang
· 5
Yao Qian
· 5
Yashesh Gaur
· 5
Zhong Meng
· 5
Zhuo Chen
· 5
Hynek Heřmanský
· 4
Midia Yousefi
· 4
Ruizhi Li
· 4
Shujie Liu
· 4
Topics
Speech Recognition
Speech Translation
Speaker Analysis
Text-to-Speech
Audio Understanding
Audio Generation
Music Generation
Multimodal Audio
Speech Enhancement