Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Wenhan Xiong — most-cited papers & profile · Multimodal
← authors
·
overview
Wenhan Xiong
14
papers ·
1441
citations ·
22
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Unified Multimodal Pre-training and Prompt-based Tuning for Vision-Language Understanding and Generation
2021 · 8 citations
VideoOFA: Two-Stage Pre-Training for Video-to-Text Generation
2023 · 2 citations
The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task
2023 · 1 citations
Coarse-to-Fine Contrastive Learning in Image-Text-Graph Space for Improved Vision-Language Compositionality
2023
Top co-authors
Pengchuan Zhang
· 2
Barlas Oguz
· 1
Barlas O\u{g}uz
· 1
Harman Singh
· 1
James C. Gee
· 1
Jingfei Du
· 1
Jingjing Chen
· 1
Lili Yu
· 1
Mengjiao Wang
· 1
Qifan Wang
· 1
Tianyi Liu
· 1
Wen-tau Yih
· 1
Topics
cs.CV
cs.CL
Video-Language
Vision-Language Models
cs.LG
Image-Text Retrieval
Benchmarks
Visual QA & Reasoning