Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Qian Chen โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Qian Chen
86
papers ยท
270
citations ยท
22
h-index
Capital University ยท Capital Medical University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
2025 ยท 86 citations
Controllable Time-Delay Transformer for Real-Time Punctuation Prediction and Disfluency Detection
2020 ยท 33 citations
LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
2023 ยท 18 citations
Pre-training for Spoken Language Understanding with Joint Textual and Phonetic Representation Learning
2021 ยท 5 citations
3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement
2023 ยท 5 citations
PoNet: Pooling Network for Efficient Token Mixing in Long Sequences
2021 ยท 4 citations
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
2024 ยท 4 citations
CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking
2023 ยท 3 citations
CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
2024 ยท 3 citations
CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
2024 ยท 3 citations
UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook
2025 ยท 2 citations
Sequence Model with Self-Adaptive Sliding Window for Efficient Spoken Document Segmentation
2021 ยท 2 citations
Improving Speaker Diarization using Semantic Information: Joint Pairwise Constraints Propagation
2023 ยท 2 citations
Adaptive Knowledge Distillation between Text and Speech Pre-trained Models
2023 ยท 1 citations
Semantic VAD: Low-Latency Voice Activity Detection for Speech Interaction
2023 ยท 1 citations
Top co-authors
Wen Wang
ยท 23
Shiliang Zhang
ยท 18
Zhihao Du
ยท 13
Qinglin Zhang
ยท 11
Yafeng Chen
ยท 9
Chong Deng
ยท 8
Hui Wang
ยท 8
Fan Yu
ยท 7
Zhifu Gao
ยท 7
Chong Zhang
ยท 6
Jiaqing Liu
ยท 6
Tianyu Zhao
ยท 6
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Audio Understanding
Multimodal Audio
cs.SD
Speech Translation
Speaker Analysis
Voice Cloning
eess.AS