Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hui Li — most-cited papers & profile · Multimodal
← authors
·
overview
Hui Li
138
papers ·
7327
citations ·
44
h-index
Central South University · Hunan University · First Affiliated Hospital of Jiangxi Medical College · 117th Hospital of People's Liberation Army · Tianjin Hospital · Second Xiangya Hospital of Central South University · Tianjin haihe hospital
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MST-Distill: Mixture of Specialized Teachers for Cross-Modal Knowledge Distillation
2025 · 1 citations
Steering Vision-Language Models with Joint Sparse Autoencoders
2026
Prompt Reinjection: Alleviating Prompt Forgetting in Multimodal Diffusion Transformers
2026
The Thinking Pixel: Recursive Sparse Reasoning In Multimodal Diffusion Latents
2026
Semantic and Visual Evidence for Efficient Long-Video Reasoning: A Solution for the HD-EPIC VQA Challenge
2026
RASR: Retrieval-Augmented Semantic Reasoning for Fake News Video Detection
2026
Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval
2025
Multipriv: Benchmarking Individual-level Privacy Reasoning In Vision-language Models
2025
HCC-3D: Hierarchical Compensatory Compression For 98% 3D Token Reduction In Vision-language Models
2025
Top co-authors
Yuxuan Yao
· 2
Huizhen Shu
· 1
Jiaming Zhang
· 1
Jingdong Wang
· 1
Jinsong Su
· 1
Jin Wang
· 1
Jun Li
· 1
Kun Zhang
· 1
Qiong Wu
· 1
Qipeng Guo
· 1
Quan Wang
· 1
Rongrong Ji
· 1
Topics
Vision-Language Models
Video-Language
Benchmarks
Visual QA & Reasoning
Instruction Tuning
Audio-Visual
Image-Text Retrieval