Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guobin Ma — most-cited papers & profile · Speech Audio
← authors
·
overview
Guobin Ma
11
papers ·
78
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DiffRhythm: Blazingly Fast and Embarrassingly Simple End-to-End Full-Length Song Generation with Latent Diffusion
2025 · 51 citations
OSUM-EChat: Enhancing End-to-End Empathetic Spoken Chatbot via Understanding-Driven Spoken Dialogue
2025 · 19 citations
MeanVC: Lightweight and Streaming Zero-Shot Voice Conversion via Mean Flows
2025 · 4 citations
LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement
2025 · 4 citations
MeanVC 2: Robust Low-Latency Streaming Zero-Shot Voice Conversion
2026
FlashTTS: Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation
2026
MINT-Bench: A Comprehensive Multilingual Benchmark for Instruction-Following Text-to-Speech
2026
OmniCodec: Low Frame Rate Universal Audio Codec with Semantic-Acoustic Disentanglement
2026
YingMusic-Singer-Plus: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance
2026
dLLM-ASR: A Faster Diffusion LLM-based Framework for Speech Recognition
2026
SynthVC: Leveraging Synthetic Data for End-to-End Low Latency Streaming Voice Conversion
2025
Top co-authors
Lei Xie
· 11
Dake Guo
· 5
Hanke Xie
· 5
Huakang Chen
· 5
Yuepeng Jiang
· 5
Jingbin Hu
· 4
Wenhao Li
· 3
Wenjie Tian
· 3
Ziqian Ning
· 3
Chengyou Wang
· 2
Chuan Xie
· 2
Chunbo Hao
· 2
Topics
cs.SD
eess.AS
Music Generation
Audio Generation
Speech Enhancement
Audio Understanding
Multimodal Audio