Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Cong Han — most-cited papers & profile · Speech Audio
← authors
·
overview
Cong Han
23
papers ·
174
citations ·
34
h-index
Shandong University of Traditional Chinese Medicine · Affiliated Hospital of Shandong University of Traditional Chinese Medicine · Henan Normal University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Real-time binaural speech separation with preserved spatial cues
2020 · 52 citations
FaSNet: Low-latency Adaptive Beamforming for Multi-microphone Audio Processing
2019 · 30 citations
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
2023 · 23 citations
StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis
2022 · 16 citations
Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions
2023 · 14 citations
Dual-path Mamba: Short and Long-term Bidirectional Selective Structured State Space Models for Speech Separation
2024 · 11 citations
Continuous Speech Separation Using Speaker Inventory for Long Multi-talker Recording
2020 · 6 citations
Online Binaural Speech Separation of Moving Speakers With a Wavesplit Network
2023 · 5 citations
Improved Decoding of Attentional Selection in Multi-Talker Environments with Self-Supervised Learned Speech Representation
2023 · 4 citations
Ultra-Lightweight Speech Separation via Group Communication
2020 · 3 citations
Rethinking the Separation Layers in Speech Separation Networks
2020 · 3 citations
SLMGAN: Exploiting Speech Language Model Representations for Unsupervised Zero-Shot Voice Conversion in GANs
2023 · 2 citations
Distortion-controlled Training for End-to-end Reverberant Speech Separation with Auxiliary Autoencoding Loss
2020 · 1 citations
StyleTTS-VC: One-Shot Voice Conversion by Knowledge Transfer from Style-Based TTS Models
2022 · 1 citations
Unsupervised Multi-channel Separation and Adaptation
2023 · 1 citations
Top co-authors
Nima Mesgarani
· 16
Yi Luo
· 7
Xilin Jiang
· 5
Zhuo Chen
· 3
Shinji Watanabe
· 2
Vishal Choudhari
· 1
Yanmin Qian
· 1
Topics
Speech Enhancement
Speech Recognition
Audio Understanding
Speech Translation
Text-to-Speech
Audio Generation
Speaker Analysis
Voice Cloning