Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jiawen Kang — most-cited papers & profile · Speech Audio
← authors
·
overview
Jiawen Kang
50
papers ·
196
citations ·
69
h-index
Guangdong University of Technology · Ministry of Education · Guangdong Institute of Intelligent Manufacturing · Ministry of Industry and Information Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
A Sidecar Separator Can Convert a Single-Talker Speech Recognition System to a Multi-Talker One
2023 · 16 citations
CN-Celeb: multi-genre speaker recognition
2020 · 9 citations
Cross-Speaker Encoding Network for Multi-Talker Speech Recognition
2024 · 4 citations
Domain-Invariant Speaker Vector Projection by Model-Agnostic Meta-Learning
2020 · 3 citations
Unified Modeling of Multi-Talker Overlapped Speech Recognition and Diarization with a Sidecar Separator
2023 · 1 citations
Large Language Model Can Transcribe Speech in Multi-Talker Scenarios with Versatile Instructions
2024
Disentangling Speakers in Multi-Talker Speech Recognition with Speaker-Aware CTC
2024
The CUHK-TENCENT speaker diarization system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge
2022
Towards Effective and Compact Contextual Representation for Conformer Transducer Speech Recognition Systems
2023
QS-TTS: Towards Semi-Supervised Text-to-Speech Synthesis via Vector-Quantized Self-Supervised Speech Representation Learning
2023
Empowering Whisper as a Joint Multi-Talker and Target-Talker Speech Recognition System
2024
Exploring SSL Discrete Speech Features for Zipformer-based Contextual ASR
2024
Top co-authors
Helen Meng
· 8
Xixin Wu
· 8
Xunying Liu
· 6
Lingwei Meng
· 5
Mingyu Cui
· 3
Yuejiao Wang
· 3
Dong Wang
· 2
Haibin Wu
· 2
Haohan Guo
· 2
Jiajun Deng
· 2
Lantian Li
· 2
Lingwei Meng
· 2
Topics
Speech Recognition
Speaker Analysis
Audio Understanding
Speech Translation
Text-to-Speech
Speech Enhancement
Audio Generation