Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Qian Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Qian Chen
25
papers ·
68
citations ·
22
h-index
Capital University · Capital Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
2023 · 18 citations
3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement
2023 · 5 citations
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
2024 · 4 citations
CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking
2023 · 3 citations
CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
2024 · 3 citations
CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
2024 · 3 citations
UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook
2025 · 2 citations
Sequence Model with Self-Adaptive Sliding Window for Efficient Spoken Document Segmentation
2021 · 2 citations
Improving Speaker Diarization using Semantic Information: Joint Pairwise Constraints Propagation
2023 · 2 citations
Adaptive Knowledge Distillation between Text and Speech Pre-trained Models
2023 · 1 citations
Semantic VAD: Low-Latency Voice Activity Detection for Speech Interaction
2023 · 1 citations
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition
2023 · 1 citations
Speech Token Prediction via Compressed-to-fine Language Modeling for Speech Generation
2025
BeamTransformer: Microphone Array-based Overlapping Speech Detection
2021
ProsoSpeech: Enhancing Prosody With Quantized Vector Pre-training in Text-to-Speech
2022
Top co-authors
Shiliang Zhang
· 10
Siqi Zheng
· 7
Zhihao Du
· 5
Wen Wang
· 4
Yafeng Chen
· 4
Zhifu Gao
· 4
Ziyang Ma
· 4
Hui Wang
· 3
Luyao Cheng
· 3
Qinglin Zhang
· 3
Shiliang Zhang
· 3
Zhijie Yan
· 3
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Audio Generation
Speech Translation
Speech Enhancement
Speaker Analysis
Multimodal Audio
Music Generation
Voice Cloning