Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xin Yu — most-cited papers & profile · Multimodal
← authors
·
overview
Xin Yu
19
papers ·
38
citations ·
10
h-index
Nvidia (United States) · Beihang University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Cluster-aware Prompt Ensemble Learning For Few-shot Vision-language Model Adaptation
2025 · 5 citations
Dynamic Orchestration Of Multi-agent System For Real-world Multi-image Agricultural VQA
2025 · 1 citations
MRD: Multi-resolution Retrieval-Detection Fusion for High-Resolution Image Understanding
2025
Top co-authors
Fan Yang
· 1
Heming Du
· 1
Kaihao Zhang
· 1
Wei Liu
· 1
Wenhan Luo
· 1
Xiaohui Tao
· 1
Xingping Dong
· 1
Yan Ke
· 1
Zhi Chen
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Video-Language
Embodied & Agents
Image-Text Retrieval