Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shansong Liu — most-cited papers & profile · Speech Audio
← authors
·
overview
Shansong Liu
11
papers ·
49
citations ·
15
h-index
Beijing Academy of Artificial Intelligence · China Telecom (China) · China Telecom
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Bayesian Transformer Language Models for Speech Recognition
2021 · 20 citations
Audio-visual Recognition of Overlapped speech for the LRS2 dataset
2020 · 10 citations
A Hierarchical Speaker Representation Framework for One-shot Singing Voice Conversion
2022 · 9 citations
Bayesian Learning of LF-MMI Trained Time Delay Neural Networks for Speech Recognition
2020 · 4 citations
Neural Architecture Search For LF-MMI Trained Time Delay Neural Networks
2020 · 2 citations
Audio-visual Multi-channel Integration and Recognition of Overlapped Speech
2020 · 2 citations
Adversarial Data Augmentation for Disordered Speech Recognition
2021 · 2 citations
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation
2026
Editing Music with Melody and Text: Using ControlNet for Diffusion Transformer
2024
Exploiting Cross Domain Acoustic-to-articulatory Inverted Features For Disordered Speech Recognition
2022
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
2024
Top co-authors
Helen Meng
· 7
Xunying Liu
· 7
Mengzhe Geng
· 6
Jianwei Yu
· 5
Shoukang Hu
· 5
Xurong Xie
· 3
Bo Wu
· 2
Dong Yu
· 2
Mingyu Cui
· 2
Shi-Xiong Zhang
· 2
Ying Shan
· 2
Zi Ye
· 2
Topics
Speech Recognition
Multimodal Audio
Speech Translation
Audio Generation
Music Generation
Speech Enhancement
Voice Cloning
Speaker Analysis
Audio Understanding