Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guanrou Yang — most-cited papers & profile · Speech Audio
← authors
·
overview
Guanrou Yang
14
papers ·
4
citations ·
5
h-index
Shanghai Jiao Tong University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
2024 · 4 citations
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling
2026
WavCube: Unifying Speech Representation for Understanding and Generation via Semantic-Acoustic Joint Modeling
2026
SemanticVocoder: Bridging Audio Generation and Audio Understanding via Semantic Latents
2026
Towards Fine-Grained and Multi-Granular Contrastive Language-Speech Pre-training
2026
DiSTAR: Diffusion over a Scalable Token Autoregressive Representation for Speech Generation
2025
CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training
2025
Speech Token Prediction via Compressed-to-fine Language Modeling for Speech Generation
2025
Pushing the Limits of Unsupervised Unit Discovery for SSL Speech Representation
2023
Fast-HuBERT: An Efficient Training Framework for Self-Supervised Speech Representation Learning
2023
MaLa-ASR: Multimedia-Assisted LLM-Based ASR
2024
TacoLM: GaTed Attention Equipped Codec Language Model are Efficient Zero-Shot Text to Speech Synthesizers
2024
Enhancing Low-Resource ASR through Versatile TTS: Bridging the Data Gap
2024
CTC-Assisted LLM-Based Contextual ASR
2024
Top co-authors
Ziyang Ma
· 7
Shiliang Zhang
· 4
Zhifu Gao
· 4
Xie Chen
· 3
Yakun Song
· 3
Zhihao Du
· 3
Fan Yu
· 2
Qian Chen
· 2
Yafeng Chen
· 2
Yushen Chen
· 2
Zhikang Niu
· 2
Zhisheng Zheng
· 2
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Audio Understanding
Multimodal Audio
Speech Translation
Music Generation
Speech Enhancement
Speaker Analysis