Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Rui Sun — most-cited papers & profile · Multimodal
← authors
·
overview
Rui Sun
4
papers ·
11
citations ·
9
h-index
Yanbian University · Jilin University · Wuhan University · Beijing Institute of Graphic Communication · Jilin Medical University · State Key Laboratory of Supramolecular Structure and Materials
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models
2023 · 4 citations
UniFine: A Unified and Fine-grained Approach for Zero-shot Vision-Language Understanding
2023 · 2 citations
JourneyBench: A Challenging One-Stop Vision-Language Understanding Benchmark of Generated Images
2024
Top co-authors
Haoxuan You
· 3
Kai-Wei Chang
· 2
Shih-Fu Chang
· 2
Zhecan Wang
· 2
Alvi Ishmam
· 1
Anushka Sivakumar
· 1
Chia-Wei Tang
· 1
Chris Thomas
· 1
Gengyu Wang
· 1
Hammad A. Ayyubi
· 1
Hammad Ayyubi
· 1
Hani Alomari
· 1
Topics
Vision-Language Models
Video-Language
Visual QA & Reasoning
Image-Text Retrieval
Benchmarks