Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yueting Zhuang — most-cited papers & profile · Multimodal
← authors
·
overview
Yueting Zhuang
60
papers ·
587
citations ·
62
h-index
Zhejiang University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Fine-Grained Semantically Aligned Vision-Language Pre-Training
2022 · 29 citations
Fine-tuning Multimodal LLMs to Follow Zero-shot Demonstrative Instructions
2023 · 11 citations
Gradient-Regulated Meta-Prompt Learning for Generalizable Vision-Language Models
2023 · 7 citations
HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation
2025 · 6 citations
BOSS: Bottom-up Cross-modal Semantic Composition with Hybrid Counterfactual Training for Robust Content-based Image Retrieval
2022 · 5 citations
Continual Vision-Language Representation Learning with Off-Diagonal Information
2023 · 2 citations
Auto-Encoding Morph-Tokens for Multimodal LLM
2024
Top co-authors
Siliang Tang
· 7
Juncheng Li
· 5
Wenqiao Zhang
· 4
Longhui Wei
· 3
Mengze Li
· 3
Qi Tian
· 3
Tat-Seng Chua
· 3
Hanwang Zhang
· 2
Kaihang Pan
· 2
Minghe Gao
· 2
Beng Chin Ooi
· 1
Binhe Yu
· 1
Topics
Vision-Language Models
Video-Language
Image-Text Retrieval
Visual QA & Reasoning
Instruction Tuning
Benchmarks