Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Wanxiang Che — most-cited papers & profile · Multimodal
← authors
·
overview
Wanxiang Che
23
papers ·
89
citations ·
49
h-index
Harbin Institute of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding
2020 · 59 citations
CVLUE: A New Benchmark Dataset for Chinese Vision-Language Understanding Evaluation
2024 · 1 citations
What Factors Affect Multi-Modal In-Context Learning? An In-Depth Exploration
2024 · 1 citations
Visual Thoughts: A Unified Perspective of Understanding Multimodal Chain-of-Thought
2025
M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought
2024
Exploring Multi-Grained Concept Annotations for Multimodal Large Language Models
2024
CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models
2024
Top co-authors
Libo Qin
· 5
Hao Fei
· 3
Qiguang Chen
· 2
Qiguang Chen
· 2
Xiao Xu
· 2
Zihui Cheng
· 2
Alex Jinpeng Wang
· 1
Cha Zhang
· 1
Chen Huang
· 1
Dinei Florencio
· 1
Fei Yu
· 1
Furu Wei
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Video-Language
Image-Text Retrieval
Audio-Visual
Instruction Tuning