Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ruibo Fu — most-cited papers & profile · Speech Audio
← authors
·
overview
Ruibo Fu
23
papers ·
11
citations ·
14
h-index
Chinese Academy of Sciences · Institute of Automation · Tsinghua University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Low-rank Adaptation Method for Wav2vec2-based Fake Audio Detection
2023 · 5 citations
Singing-Tacotron: Global duration control attention and dynamic filter for End-to-end singing voice synthesis
2022 · 3 citations
Half-Truth: A Partially Fake Audio Detection Dataset
2021 · 1 citations
UnifySpeech: A Unified Framework for Zero-shot Text-to-Speech and Voice Conversion
2023 · 1 citations
Generalized Fake Audio Detection via Deep Stable Learning
2024 · 1 citations
M3-TTS: Multi-modal DiT Alignment & Mel-latent for Zero-shot High-fidelity Speech Synthesis
2025
VQ-CTAP: Cross-Modal Fine-Grained Sequence Representation Learning for Speech Processing
2024
CampNet: Context-Aware Mask Prediction for End-to-End Text-Based Speech Editing
2022
Fully Automated End-to-End Fake Audio Detection
2022
Text Enhancement for Paragraph Processing in End-to-End Code-switching TTS
2022
Emotion Selectable End-to-End Text-based Speech Editing
2022
Minimally-Supervised Speech Synthesis with Conditional Diffusion Model and Language Model: A Comparative Study of Semantic Coding
2023
Learning Speech Representation From Contrastive Token-Acoustic Pretraining
2023
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection
2023
Codecfake: An Initial Dataset for Detecting LLM-based Deepfake Audio
2024
Top co-authors
Zhengqi Wen
· 13
Jianhua Tao
· 12
Jiangyan Yi
· 10
Chunyu Qiang
· 8
Tao Wang
· 7
Yuankun Xie
· 7
Shuchen Shi
· 6
Chenxing Li
· 4
Xin Qi
· 4
Yukun Liu
· 4
Guanjun Li
· 3
Jianwu Dang
· 3
Topics
Audio Generation
Text-to-Speech
Speech Recognition
Speech Enhancement
Audio Understanding
Multimodal Audio
Speaker Analysis
Voice Cloning
Speech Translation
Music Generation