Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhengyang Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Zhengyang Chen
13
papers ·
73
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Self-Supervised Speaker Verification Using Dynamic Loss-Gate and Label Correction
2022 · 29 citations
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
2024 · 24 citations
UniSpeech-SAT: Universal Speech Representation Learning with Speaker Aware Pre-Training
2021 · 9 citations
Large-scale Self-Supervised Speech Representation Learning for Automatic Speaker Verification
2021 · 9 citations
Attention-based Encoder-Decoder End-to-End Neural Diarization with Embedding Enhancer
2023 · 1 citations
Target Speech Diarization with Multimodal Prompts
2024 · 1 citations
Training Text-to-Speech Model with Purely Synthetic Data: Feasibility, Sensitivity, and Generalization Capability
2025
The SJTU System for Short-duration Speaker Verification Challenge 2021
2022
Attention-based Encoder-Decoder Network for End-to-End Neural Speaker Diarization with Target Speaker Attractor
2023
Prompt-driven Target Speech Diarization
2023
Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
2024
Flow-TSVAD: Target-Speaker Voice Activity Detection via Latent Flow Matching
2024
Disentangling the Prosody and Semantic Information with Pre-trained Model for In-Context Learning based Zero-Shot Voice Conversion
2024
Top co-authors
Yanmin Qian
· 11
Bing Han
· 4
Haizhou Li
· 3
Yidi Jiang
· 3
Chengyi Wang
· 2
Junichi Yamagishi
· 2
Ruijie Tao
· 2
Sanyuan Chen
· 2
Shuai Wang
· 2
Shujie Liu
· 2
Xuechen Liu
· 2
Yao Qian
· 2
Topics
Speaker Analysis
Speech Recognition
Audio Understanding
Text-to-Speech
Audio Generation
Speech Translation
Voice Cloning
Speech Enhancement
Multimodal Audio