Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chong Deng — most-cited papers & profile · Speech Audio
← authors
·
overview
Chong Deng
12
papers ·
89
citations ·
12
h-index
Hong Kong Science and Technology Parks Corporation · City University of Hong Kong, Shenzhen Research Institute · Beihang University · University of Hong Kong
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
2025 · 86 citations
CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
2024 · 3 citations
Fun-Audio-Chat Technical Report
2025
Say More with Less: Variable-Frame-Rate Speech Tokenization via Adaptive Clustering and Implicit Duration Coding
2025
Fun-ASR Technical Report
2025
SpeakerLM: End-to-End Versatile Speaker Diarization and Recognition with Multimodal Large Language Models
2025
DrVoice: Parallel Speech-Text Voice Conversation Model via Dual-Resolution Speech Representations
2025
OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
2024
Loss Masking Is Not Needed in Decoder-only Transformer for Discrete-token-based ASR
2023
Top co-authors
Wen Wang
· 9
Qian Chen
· 8
Qinglin Zhang
· 7
Jiaqing Liu
· 5
Hui Wang
· 4
Shiliang Zhang
· 4
Xiangang Li
· 4
Xiang Lv
· 4
Zhihao Du
· 4
Changfeng Gao
· 3
Chong Zhang
· 3
Jieping Ye
· 3
Topics
Audio Generation
cs.SD
Multimodal Audio
Text-to-Speech
Speech Recognition
cs.AI
eess.AS
Voice Cloning
cs.CL
Speech Enhancement