Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guangzhi Sun — most-cited papers & profile · Multimodal
← authors
·
overview
Guangzhi Sun
5
papers ·
34
citations ·
40
h-index
University of Cambridge · Shandong Provincial QianFoShan Hospital · Shandong First Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SALMONN-omni: A Standalone Speech LLM without Codec Injection for Full-duplex Conversation
2025 · 16 citations
Parameter Efficient Finetuning for Speech Emotion Recognition and Domain Adaptation
2024 · 10 citations
Low-Rank and Sparse Model Merging for Multi-Lingual Speech Recognition and Translation
2025 · 5 citations
Audio-Conditioned Diffusion LLMs for ASR and Deliberation Processing
2025 · 3 citations
Cross-Lingual Interleaving for Speech Language Models
2025
Topics
Speech Recognition
Audio Understanding
Speech Translation
Multimodal Audio
Speech Enhancement
Speaker Analysis