Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ruibin Yuan — most-cited papers & profile · Speech Audio
← authors
·
overview
Ruibin Yuan
17
papers ·
86
citations ·
8
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
2023 · 27 citations
CLaMP 3: Universal Music Information Retrieval Across Unaligned Modalities and Unseen Languages
2025 · 5 citations
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
2023 · 3 citations
MAP-Music2Vec: A Simple and Effective Baseline for Self-Supervised Music Audio Representation Learning
2022 · 2 citations
On the Effectiveness of Speech Self-supervised Learning for Music
2023 · 2 citations
Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
2025 · 1 citations
WenetSpeech-Yue: A Large-scale Cantonese Speech Corpus with Multi-dimensional Annotation
2025
UniSS: Unified Expressive Speech-to-Speech Translation with Your Voice
2025
AudioX: A Unified Framework for Anything-to-Audio Generation
2025
Editing Music with Melody and Text: Using ControlNet for Diffusion Transformer
2024
SongTrans: An unified song transcription and alignment method for lyrics and notes
2024
Top co-authors
Yike Guo
· 6
Chenghua Lin
· 5
Emmanouil Benetos
· 4
Yinghao Ma
· 4
Anton Ragni
· 3
Gus Xia
· 3
Hanzhi Yin
· 3
Jie Fu
· 3
Roger Dannenberg
· 3
Ruibo Liu
· 3
Xingran Chen
· 3
Yizhi Li
· 3
Topics
Music Generation
Audio Understanding
Multimodal Audio
Audio Generation
Speech Recognition
Text-to-Speech
Voice Cloning
Speaker Analysis
Speech Translation