Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yafeng Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Yafeng Chen
15
papers ·
102
citations ·
40
h-index
Jiangyin Traffic Planning Survey & Design Institute (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
2025 · 86 citations
3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement
2023 · 5 citations
Improved Meta-Learning Training for Speaker Verification
2021 · 4 citations
CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking
2023 · 3 citations
Exploring Universal Speech Attributes for Speaker Verification with an Improved Cross-stitch Network
2020 · 2 citations
Improving Speaker Diarization using Semantic Information: Joint Pairwise Constraints Propagation
2023 · 2 citations
PerTTS: Personalized and Controllable Zero-Shot Spontaneous Style Text-to-Speech Synthesis
2026
SpeakerLM: End-to-End Versatile Speaker Diarization and Recognition with Multimodal Large Language Models
2025
Exploring Efficient Directional and Distance Cues for Regional Speech Separation
2025
DrVoice: Parallel Speech-Text Voice Conversation Model via Dual-Resolution Speech Representations
2025
CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training
2025
Speech Token Prediction via Compressed-to-fine Language Modeling for Speech Generation
2025
Multi-task Metric Learning for Text-independent Speaker Verification
2020
Graph Convolutional Network Based Semi-Supervised Learning on Multi-Speaker Meeting Data
2022
Top co-authors
Qian Chen
· 9
Hui Wang
· 7
Wen Wang
· 5
Chong Deng
· 3
Guanrou Yang
· 3
Qinglin Zhang
· 3
Shiliang Zhang
· 3
Tianyu Zhao
· 3
Wu Guo
· 3
Xiang Lv
· 3
Changfeng Gao
· 2
Chongjia Ni
· 2
Topics
Speaker Analysis
Speech Recognition
Text-to-Speech
Audio Generation
Voice Cloning
Audio Understanding
Multimodal Audio
cs.SD
Speech Enhancement
Speech Translation