Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xin Huang — most-cited papers & profile · Multimodal
← authors
·
overview
Xin Huang
1
papers ·
0
citations ·
30
h-index
Hong Kong Baptist University · Shandong Institute of Business and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
M-DocSum: Do LVLMs Genuinely Comprehend Interleaved Image-Text in Document Summarization?
2025
Top co-authors
Daxin Jiang
· 1
Haolong Yan
· 1
Kaijun Tan
· 1
Si Li
· 1
Xiangyu Zhang
· 1
Yeqing Shen
· 1
Zheng Ge
· 1
Topics
Image-Text Retrieval
Vision-Language Models
Benchmarks