Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shuo Yang — most-cited papers & profile · Multimodal
← authors
·
overview
Shuo Yang
114
papers ·
750
citations ·
16
h-index
Sun Yat-sen University · Fudan University · Zhongshan Hospital · The First Affiliated Hospital, Sun Yat-sen University · Shandong Agricultural University · Guangzhou Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
BiCro: Noisy Correspondence Rectification for Multi-modality Data via Bi-directional Cross-modal Similarity Consistency
2023 · 3 citations
Robofac: A Comprehensive Framework For Robotic Failure Analysis And Correction
2025
Attention in Space: Functional Roles of VLM Heads for Spatial Reasoning
2026
Bridging Perception and Reasoning: Token Reweighting for RLVR in Multimodal LLMs
2026
Beyond Where to Look: Trajectory-Guided Reinforcement Learning for Multimodal RLVR
2026
Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference
2025
GNSP: Gradient Null Space Projection for Preserving Cross-Modal Alignment in VLMs Continual Learning
2025
LLM-enhanced Action-aware Multi-modal Prompt Tuning for Image-Text Matching
2025
Logic Unseen: Revealing The Logical Blindspots Of Vision-language Models
2025
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
2025
Top co-authors
Bo Zhao
· 2
Jiancan Wu
· 2
Jinda Lu
· 2
Jinghan Li
· 2
Junkang Wu
· 2
Kexin Huang
· 2
Xiang Wang
· 2
Eduard Hovy
· 1
Guoyin Wang
· 1
Hongxun Yao
· 1
Hongyi Cai
· 1
Kai Wang
· 1
Topics
Vision-Language Models
Benchmarks
Visual QA & Reasoning
Image-Text Retrieval
Embodied & Agents
Audio-Visual
Video-Language