Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Luowei Zhou — most-cited papers & profile · Multimodal
← authors
·
overview
Luowei Zhou
7
papers ·
193
citations ·
38
h-index
Bellevue University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Unified Vision-Language Pre-Training for Image Captioning and VQA
2019 · 74 citations
CLIP-TD: CLIP Targeted Distillation for Vision-Language Tasks
2022 · 19 citations
AssistGPT: A General Multi-modal Assistant that can Plan, Execute, Inspect, and Learn
2023 · 19 citations
Multimodal Adaptive Distillation for Leveraging Unimodal Encoders for Vision-Language Tasks
2022 · 8 citations
UC2: Universal Cross-lingual Cross-modal Vision-and-Language Pre-training
2021 · 7 citations
MIST: Multi-modal Iterative Spatial-Temporal Transformer for Long-form Video Question Answering
2022 · 5 citations
Top co-authors
Bin Xiao
· 2
Difei Gao
· 2
Haoxuan You
· 2
Jianwei Yang
· 2
Lu Yuan
· 2
Mike Zheng Shou
· 2
Noel Codella
· 2
Shih-Fu Chang
· 2
Xiyang Dai
· 2
Yen-Chun Chen
· 2
Zhecan Wang
· 2
Hamid Palangi
· 1
Topics
Video-Language
Visual QA & Reasoning
Vision-Language Models
Benchmarks
Image-Text Retrieval